Qwen3 VL 4B (Reasoning)

Qwen3 VL 4B (Reasoning)

Alibaba · Released Oct 14, 2025
Intelligence #717 / 803
13.2 our score
AA Index #334 / 444
7.0 Artificial Analysis
Input Price
Not priced
Output Price
Not priced
Context
Not reported

Analysis Summary

Qwen3 VL 4B Reasoning is Alibaba's compact vision-language reasoning model, with supplied measurements across mathematics, general reasoning, knowledge, coding, instruction following, long-context work, terminal tasks, and tool-oriented evaluation. Its instruction-following result is stronger than the non-reasoning variant, while terminal performance is very low.

That profile suits image-aware structured content, document extraction, and lightweight analytical assistance more than autonomous coding or complex agents. The model may be useful where compact multimodal processing matters, but no pricing, context-window size, throughput, or API details are supplied. Human review remains important for client-facing outputs and tool-driven actions.

Assessed September 7, 2026

Editorial notes

Qwen3 VL 4B Reasoning adds stronger instruction, context, reasoning, coding, and tool-oriented measurements to a compact multimodal model, but its low terminal performance limits autonomous engineering use.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Qwen3 VL 4B Reasoning adds stronger instruction, context, reasoning, coding, and tool-oriented measurements to a compact multimodal model, but its low terminal performance limits autonomous engineering use.

#717 of 803 overall Down 36 this week

Benchmark scores

GPQA Diamond 49.4%
HLE 4.4%
MMLU Pro 70%
AIME 2025 25.7%
SciCode 17.1%
LiveCodeBench 32%
TerminalBench Hard 1.5%
τ²-Bench 15.5%
IFBench 36.6%
LCR 21.3%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

7 Intelligence Index·25.7 Math Index

How Qwen3 VL 4B (Reasoning) compares

Qwen3 VL 4B (Reasoning) ranks #334 of 444 AI models we track for overall intelligence. Qwen3 VL 4B (Reasoning) is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 11% of models #717
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Qwen3 VL 4B (Reasoning) $0.00 in $0.00 out
Claude Opus 5 $5.00 in $25.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out
Qwen: Qwen3.8 2.4T A95B $2.00 in $6.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence2Technical0Value0Content3.4
Performance profile

Strongest on content. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Frequently asked questions about Qwen3 VL 4B (Reasoning)

How much does Qwen3 VL 4B (Reasoning) cost?

Qwen3 VL 4B (Reasoning) is currently available for free via OpenRouter.

Who created Qwen3 VL 4B (Reasoning)?

Qwen3 VL 4B (Reasoning) is developed by Alibaba and was released on October 14, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.