Gemma 4 12B (Reasoning)

Gemma 4 12B (Reasoning)

Google · Released Jun 3, 2026
Intelligence #320 / 688
32.7 our score
Speed #132 / 314
111.0 tok/s
Input Price
Not priced
Output Price
Not priced
Context
Not reported

Analysis Summary

Gemma 4 12B Reasoning is a compact Google model with measured gains in reasoning, coding, and instruction following relative to a basic lightweight model. The supplied data does not list context capacity, pricing, tool use, function calling, or multimodal support, so its deployment profile is narrower than models designed explicitly for agents. Terminal and agentic results are weak for autonomous execution.

It can serve supervised drafting, structured extraction, basic code assistance, and internal analysis where a human checks the result. Its instruction-following performance is useful for templates and controlled prompts, but long-context document work and complex tool chains are not supported by the available evidence. Adopt it as a lightweight task model after local validation, not as the primary engine for client-facing content or software agents.

Assessed August 9, 2026

Editorial notes

Gemma 4 12B Reasoning improves coding and instruction-following over the non-reasoning variant, but weak agentic and terminal results limit it to supervised, lower-risk workflows.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Gemma 4 12B Reasoning improves coding and instruction-following over the non-reasoning variant, but weak agentic and terminal results limit it to supervised, lower-risk workflows.

#320 of 688 overall Down 1 this week

Benchmark scores

GPQA Diamond 75.3%
HLE 14.8%
SciCode 38.2%
TerminalBench Hard 18.2%
τ²-Bench 36.3%
IFBench 73.5%
LCR 55.3%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

22.3 Intelligence Index·31 Coding Index·7.9 Agentic Index

How Gemma 4 12B (Reasoning) compares

Gemma 4 12B (Reasoning) ranks #156 of 420 AI models we track for overall intelligence, #104 of 193 for coding, #117 of 175 for agentic tasks. Gemma 4 12B (Reasoning) is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 53% of models #320
Speed: faster than 58% of models #132
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Gemma 4 12B (Reasoning) $0.00 in $0.00 out
Claude Opus 5 $5.00 in $25.00 out
Qwen: Qwen3.8 Max $2.00 in $6.00 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Gemma 4 12B (Reasoning) 0

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence3.7Technical2.6Value0Content7
Performance profile

Strongest on content. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Frequently asked questions about Gemma 4 12B (Reasoning)

How much does Gemma 4 12B (Reasoning) cost?

Gemma 4 12B (Reasoning) is currently available for free via OpenRouter.

Is Gemma 4 12B (Reasoning) good for coding?

On our coding benchmark index, Gemma 4 12B (Reasoning) ranks #104 of 193 models, placing it in the broader range of the field for code generation and debugging.

Who created Gemma 4 12B (Reasoning)?

Gemma 4 12B (Reasoning) is developed by Google and was released on June 3, 2026.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.