Qwen3 VL 4B (Reasoning)

Qwen3 VL 4B (Reasoning)

Alibaba · Released Oct 14, 2025
Intelligence #641 / 743
16.6 our score
AA Index #321 / 433
7.7 Artificial Analysis
Input Price
Not priced
Output Price
Not priced
Context
Not reported

Analysis Summary

Qwen3 VL 4B Reasoning is Alibaba's compact multimodal reasoning model. It has measurable capability across mathematics, knowledge, coding, instruction following, long-context tasks, and image-related use, with better instruction-following results than the corresponding non-reasoning variant.

The model fits visual document triage, image-aware content assistance, metadata generation, and simple structured extraction. Its terminal and agentic results are limited, so it should not operate unsupervised tools or manage complex multi-step workflows. No pricing or context capacity is provided in the record.

Use it when multimodal input is essential and workloads can be reviewed or constrained. For autonomous agents and demanding client content, a larger model remains the safer choice.

Assessed August 9, 2026

Editorial notes

Qwen3 VL 4B Reasoning combines compact vision support with stronger instruction following than the non-reasoning variant, but limited reasoning depth and weak terminal performance restrict it to lightweight multimodal workflows.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Qwen3 VL 4B Reasoning combines compact vision support with stronger instruction following than the non-reasoning variant, but limited reasoning depth and weak terminal performance restrict it to lightweight multimodal workflows.

#641 of 743 overall Down 25 this week

Benchmark scores

GPQA Diamond 49.4%
HLE 4.4%
MMLU Pro 70%
AIME 2025 25.7%
SciCode 17.1%
LiveCodeBench 32%
TerminalBench Hard 1.5%
τ²-Bench 15.5%
IFBench 36.6%
LCR 21.3%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

7.7 Intelligence Index·25.7 Math Index

How Qwen3 VL 4B (Reasoning) compares

Qwen3 VL 4B (Reasoning) ranks #321 of 433 AI models we track for overall intelligence. Qwen3 VL 4B (Reasoning) is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 14% of models #641
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Qwen3 VL 4B (Reasoning) $0.00 in $0.00 out
Claude Opus 5 $5.00 in $25.00 out
OpenAI: GPT-5.6 Sol $2.00 in $10.00 out
SpaceXAI: Grok 4.6 $2.00 in $6.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Qwen3 VL 4B (Reasoning) 0

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence2Technical0Value0Content5
Performance profile

Strongest on content. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Frequently asked questions about Qwen3 VL 4B (Reasoning)

How much does Qwen3 VL 4B (Reasoning) cost?

Qwen3 VL 4B (Reasoning) is currently available for free via OpenRouter.

Who created Qwen3 VL 4B (Reasoning)?

Qwen3 VL 4B (Reasoning) is developed by Alibaba and was released on October 14, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.