Qwen3.5 4B (Non-reasoning)

Qwen3.5 4B (Non-reasoning)

Alibaba · Released Mar 2, 2026
Intelligence #198 / 650
40.6 our score
Speed #290 / 296
24.1 tok/s
Input Price #146 / 688
$0.030 per 1M tokens
Output Price #169 / 688
$0.150 per 1M tokens
Context
Not reported

Analysis Summary

This non-reasoning variant of Qwen3.5 4B trims further capability from an already small model, scoring low on intelligence and coding benchmarks.

It is best reserved for lightweight, high-volume tasks like tagging or short-form generation where cost per call outweighs quality requirements.

Not recommended for coding, agentic, or client-facing content work.

Assessed July 29, 2026

Editorial notes

Qwen3.5 4B (Non-reasoning) is a compact, inexpensive model with limited reasoning and coding capability suited to simple tasks only.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Qwen3.5 4B (Non-reasoning) is a compact, inexpensive model with limited reasoning and coding capability suited to simple tasks only.

#198 of 650 overall

How Qwen3.5 4B (Non-reasoning) compares

Qwen3.5 4B (Non-reasoning) ranks #179 of 401 AI models we track for overall intelligence, #118 of 176 for coding. At $0.03 per million input tokens it is cheaper than 79% of comparable models.

Position in the field
Intelligence: smarter than 70% of models #198
Speed: faster than 2% of models #290
Price: cheaper than 79% of models #146
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Qwen3.5 4B (Non-reasoning) $0.03 in $0.15 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out
Claude Opus 5 $5.00 in $25.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence2.3Technical3.6Value7.3Content3
Performance profile

Strongest on value. The pulled-in intelligence corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $0.03 $0.000030
Output $0.15 $0.000150

What would Qwen3.5 4B (Non-reasoning) cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$0/mo Qwen3.5 4B (Non-reasoning)

Full calculator with 688 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Qwen3.5 4B (Non-reasoning) for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

Frequently asked questions about Qwen3.5 4B (Non-reasoning)

How much does Qwen3.5 4B (Non-reasoning) cost?

Qwen3.5 4B (Non-reasoning) costs $0.03 per million input tokens and $0.15 per million output tokens.

Is Qwen3.5 4B (Non-reasoning) good for coding?

On our coding benchmark index, Qwen3.5 4B (Non-reasoning) ranks #118 of 176 models, placing it in the broader range of the field for code generation and debugging.

Who created Qwen3.5 4B (Non-reasoning)?

Qwen3.5 4B (Non-reasoning) is developed by Alibaba and was released on March 2, 2026.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.