Qwen: Qwen3 4B (free)

Qwen: Qwen3 4B (free)

qwen · Released Apr 30, 2025
Intelligence #587 / 650
9.9 our score
Speed #132 / 294
105.0 tok/s
Input Price
Not priced
Output Price
Not priced
Context #464 / 654
40,960 tokens

Analysis Summary

Qwen3 4B is a small open model from Alibaba offered free of charge, with tool use and function calling over a 40K context window. Its reasoning, knowledge, and coding scores are all modest, reflecting its very small size.

It is suitable for basic chat, simple text transformation, or experimentation where budget is the primary constraint, but it should not be trusted with complex reasoning, coding, or client-facing writing tasks.

Being free makes it attractive for testing and prototyping, though businesses should expect to graduate to a stronger model for production use.

Assessed July 25, 2026

Editorial notes

Qwen3 4B is a free, lightweight model with limited reasoning and coding ability, best suited to simple, high-volume text tasks.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Qwen3 4B is a free, lightweight model with limited reasoning and coding ability, best suited to simple, high-volume text tasks.

#587 of 650 overall Down 28 this week

Benchmark scores

GPQA Diamond 39.8%
HLE 3.7%
MMLU Pro 58.6%
MATH 500 84.3%
AIME 21.3%
SciCode 16.7%
LiveCodeBench 23.3%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

6.8 Intelligence Index

How Qwen: Qwen3 4B (free) compares

Qwen: Qwen3 4B (free) ranks #303 of 398 AI models we track for overall intelligence. Its 41K-token context window is larger than 29% of the models we list. Qwen: Qwen3 4B (free) is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 10% of models #587
Speed: faster than 55% of models #132
Context: larger than 29% of models #464
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Qwen: Qwen3 4B (free) $0.00 in $0.00 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out
Claude Opus 5 $5.00 in $25.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence1.7Technical0Value0Content3
Performance profile

Strongest on content. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

About Qwen: Qwen3 4B (free)

Qwen3-4B is a 4 billion parameter dense language model from the Qwen3 series, designed to support both general-purpose and reasoning-intensive tasks. It introduces a dual-mode architecture—thinking and non-thinking—allowing dynamic switching between high-precision logical reasoning and efficient dialogue generation. This makes it well-suited for multi-turn chat, instruction following, and complex agent workflows.

Frequently asked questions about Qwen: Qwen3 4B (free)

How much does Qwen: Qwen3 4B (free) cost?

Qwen: Qwen3 4B (free) is currently available for free via OpenRouter.

What is the context window of Qwen: Qwen3 4B (free)?

Qwen: Qwen3 4B (free) has a context window of 40,960 tokens (41K).

What can Qwen: Qwen3 4B (free) do?

Qwen: Qwen3 4B (free) supports tool use and function calling.

Who created Qwen: Qwen3 4B (free)?

Qwen: Qwen3 4B (free) is developed by Qwen and was released on April 30, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.