Qwen3 4B 2507 Instruct

Qwen3 4B 2507 Instruct

Alibaba · Released Aug 6, 2025
Intelligence #552 / 650
16.1 our score
AA Index #297 / 401
7.1 Artificial Analysis
Input Price
Not priced
Output Price
Not priced
Context
Not reported

Analysis Summary

Qwen3 4B 2507 Instruct is a small model that shows reasonable results on math and general knowledge tests relative to its size, though overall reasoning capability remains modest compared to larger frontier models. Coding and agentic scores are limited.

It suits businesses running simple automation, short-form content, or basic Q&A where cost and speed outweigh the need for deep reasoning. It is not a fit for coding-heavy workflows, long-document analysis, or autonomous agent tasks.

As a small, efficient model it likely runs cheaply and quickly, making it a reasonable choice for high-volume, low-complexity use cases rather than flagship business workloads.

Assessed July 25, 2026

Editorial notes

Qwen3 4B punches above its size on math and knowledge tests but remains a lightweight model best suited to simple, cost-sensitive tasks.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Qwen3 4B punches above its size on math and knowledge tests but remains a lightweight model best suited to simple, cost-sensitive tasks.

#552 of 650 overall

Benchmark scores

GPQA Diamond 51.7%
HLE 4.7%
MMLU Pro 67.2%
AIME 2025 52.3%
SciCode 18.1%
LiveCodeBench 37.7%
TerminalBench Hard 4.5%
τ²-Bench 26.6%
IFBench 33.5%
LCR 7.3%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

7.1 Intelligence Index·52.3 Math Index

How Qwen3 4B 2507 Instruct compares

Qwen3 4B 2507 Instruct ranks #297 of 401 AI models we track for overall intelligence. Qwen3 4B 2507 Instruct is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 15% of models #552
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Qwen3 4B 2507 Instruct $0.00 in $0.00 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out
Claude Opus 5 $5.00 in $25.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence2Technical1.8Value0Content3
Performance profile

Strongest on content. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Frequently asked questions about Qwen3 4B 2507 Instruct

How much does Qwen3 4B 2507 Instruct cost?

Qwen3 4B 2507 Instruct is currently available for free via OpenRouter.

Who created Qwen3 4B 2507 Instruct?

Qwen3 4B 2507 Instruct is developed by Alibaba and was released on August 6, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.