xAI: Grok 4.20

x-ai · Released Mar 31, 2026
DFO score #75 / 697
42.4 leaderboard score
Speed #151 / 332
103.3 tok/s
Input Price #551 / 697
$1.25 per 1M tokens
Output Price #497 / 697
$2.50 per 1M tokens
Context #1 / 697
2M tokens

Analysis Summary

xAI's Grok 4.20 offers a two-million-token context window, image and file input, vision, tool use, and function calling. Its measured intelligence index is below the leading models in the supplied landscape, and its long-context result is weak despite the large advertised context. Listed rates are $1.25 per million input tokens and $2.50 per million output tokens.

The combination of multimodal input and tools may suit supervised file analysis and workflows that need substantial input capacity. Yet the measured profile does not establish dependable long-context reasoning or a strong coding specialty. Premium pricing makes it harder to justify for routine volume, particularly when lower-cost models can be evaluated for simpler tasks.

Consider it for targeted trials where the two-million-token context or multimodal inputs are central. Validate retrieval and reasoning over long files before assigning it document-heavy client work.

Assessed September 23, 2026

Editorial notes

Grok 4.20 combines a two-million-token context, vision, file input, and function calling; its measured reasoning and long-context results are comparatively limited for a premium-priced model.

The DFO score ranks models on the Artificial Analysis Intelligence Index and is refreshed as new models launch. The write-up and verdict are our own. Feedback?

DFO Verdict

Grok 4.20 combines a two-million-token context, vision, file input, and function calling; its measured reasoning and long-context results are comparatively limited for a premium-priced model.

#75 of 697 overall Down 2 this week

Benchmarks

Intelligence Index 25.7 #74 of 435 · Artificial Analysis
Coding Index Not published yet Artificial Analysis
Agentic Index Not published yet Artificial Analysis
Creative Writing Not rated Arena rating
Instruction Following Not rated Arena rating
GPQA Diamond 77.6%
HLE 24.2%
SciCode 32.8%
TerminalBench Hard 16.7%
τ²-Bench 59.9%
IFBench 49.3%
LCR 17.3%

Artificial Analysis data refreshed Oct 8, 2026. Arena ratings from the Oct 2, 2026 text leaderboard, with style control. Bars: magenta = reasoning, ink = coding and agents, cyan = instructions and long context.

Artificial Analysis has not published coding or agentic results for xAI: Grok 4.20 yet. Until it does, the model sits below tested models on the Coding and AI Agents lists.

How xAI: Grok 4.20 compares

XAI: Grok 4.20 ranks #74 of 435 AI models we track for overall intelligence. Its 2M-token context window is larger than 100% of the models we list. At $1.25 per million input tokens it is cheaper than 21% of comparable models.

Position in the field
DFO score: ahead of 89% of models #74
Speed: faster than 55% of models #151
Price: cheaper than 21% of models #551
Context: larger than 100% of models #1
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
xAI: Grok 4.20 $1.25 in $2.50 out
Anthropic: Claude Opus 5.5 $4.00 in $20.00 out
Anthropic: Claude Sonnet 5.5 $2.00 in $10.00 out
Anthropic: Claude Fable 5.1 $10.00 in $50.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $1.25 $0.001250
Output $2.50 $0.002500

What would xAI: Grok 4.20 cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$5.85/mo xAI: Grok 4.20 $0.0020 per request
$31.68/mo Anthropic: Claude Opus 5.5 $0.01 per request
$15.84/mo Anthropic: Claude Sonnet 5.5 · best value $0.0053 per request

Full calculator with 697 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save xAI: Grok 4.20 for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

About xAI: Grok 4.20

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering..

Frequently asked questions about xAI: Grok 4.20

How much does xAI: Grok 4.20 cost?

xAI: Grok 4.20 costs $1.25 per million input tokens and $2.50 per million output tokens.

What is the context window of xAI: Grok 4.20?

xAI: Grok 4.20 has a context window of 2,000,000 tokens (2M).

What can xAI: Grok 4.20 do?

xAI: Grok 4.20 supports image/vision input, tool use, and function calling.

Who created xAI: Grok 4.20?

xAI: Grok 4.20 is developed by xAI and was released on March 31, 2026.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.