Google: Gemma 4 31B

Google: Gemma 4 31B

google · Released Apr 2, 2026
Intelligence #43 / 650
68.3 our score
Speed #267 / 294
35.2 tok/s
Input Price #203 / 654
$0.100 per 1M tokens
Output Price #232 / 654
$0.340 per 1M tokens
Context #162 / 654
262,144 tokens

Analysis Summary

Gemma 4 31B is Google's dense 31-billion-parameter model, with an intelligence index of 29.4 and a coding index of 43.4, both meaningfully above the sparse 26B A4B sibling. The agentic index of 48.2 is more capable for multi-step tasks, and the model supports vision, video, tool use, and function calling across a 262K context window.

For businesses, Gemma 4 31B suits structured content generation, SEO workflows, coding assistance for lighter tasks, and tool-calling pipelines where multimodal input is needed. The instruction following score of 0.756 is strong for its tier, making it reliable for templated and structured output tasks.

At $0.12 input and $0.35 output per million tokens, it offers excellent price-performance for a benchmarked multimodal model. Teams needing a step up from the 26B A4B without moving to premium pricing will find it a practical choice.

Assessed July 10, 2026

Editorial notes

Gemma 4 31B from Google pairs vision, video, and tool use with a 262K context at $0.12 input per million tokens, offering a meaningful step up in reasoning and coding over the 26B A4B variant.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Gemma 4 31B from Google pairs vision, video, and tool use with a 262K context at $0.12 input per million tokens, offering a meaningful step up in reasoning and coding over the 26B A4B variant.

#43 of 650 overall Up 2 this week

Benchmark scores

GPQA Diamond 85.7%
HLE 22.7%
SciCode 43.4%
TerminalBench Hard 36.4%
τ²-Bench 59.9%
IFBench 75.6%
LCR 62%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

29.4 Intelligence Index·43.4 Coding Index·14.4 Agentic Index

How Google: Gemma 4 31B compares

Google: Gemma 4 31B ranks #100 of 398 AI models we track for overall intelligence, #61 of 173 for coding, #84 of 154 for agentic tasks. Its 262K-token context window is larger than 75% of the models we list. At $0.10 per million input tokens it is cheaper than 69% of comparable models.

Position in the field
Intelligence: smarter than 93% of models #43
Speed: faster than 9% of models #267
Price: cheaper than 69% of models #203
Context: larger than 75% of models #162
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Google: Gemma 4 31B $0.10 in $0.34 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out
Claude Opus 5 $5.00 in $25.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence4.8Technical5.6Value8Content6.5
Performance profile

Strongest on value. The pulled-in intelligence corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $0.10 $0.000100
Output $0.34 $0.000340

What would Google: Gemma 4 31B cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

Full calculator with 654 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Google: Gemma 4 31B for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

About Google: Gemma 4 31B

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function..

Frequently asked questions about Google: Gemma 4 31B

How much does Google: Gemma 4 31B cost?

Google: Gemma 4 31B costs $0.10 per million input tokens and $0.34 per million output tokens.

What is the context window of Google: Gemma 4 31B?

Google: Gemma 4 31B has a context window of 262,144 tokens (262K).

Is Google: Gemma 4 31B good for coding?

On our coding benchmark index, Google: Gemma 4 31B ranks #61 of 173 models, placing it in the broader range of the field for code generation and debugging.

What can Google: Gemma 4 31B do?

Google: Gemma 4 31B supports image/vision input, tool use, and function calling.

Who created Google: Gemma 4 31B?

Google: Gemma 4 31B is developed by Google and was released on April 2, 2026.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.