Analysis Summary
xAI's Grok 4.20 offers a two-million-token context window, image and file input, vision, tool use, and function calling. Its measured intelligence index is below the leading models in the supplied landscape, and its long-context result is weak despite the large advertised context. Listed rates are $1.25 per million input tokens and $2.50 per million output tokens.
The combination of multimodal input and tools may suit supervised file analysis and workflows that need substantial input capacity. Yet the measured profile does not establish dependable long-context reasoning or a strong coding specialty. Premium pricing makes it harder to justify for routine volume, particularly when lower-cost models can be evaluated for simpler tasks.
Consider it for targeted trials where the two-million-token context or multimodal inputs are central. Validate retrieval and reasoning over long files before assigning it document-heavy client work.
Assessed September 23, 2026
Editorial notes
Grok 4.20 combines a two-million-token context, vision, file input, and function calling; its measured reasoning and long-context results are comparatively limited for a premium-priced model.
The DFO score ranks models on the Artificial Analysis Intelligence Index and is refreshed as new models launch. The write-up and verdict are our own. Feedback?
DFO Verdict
Grok 4.20 combines a two-million-token context, vision, file input, and function calling; its measured reasoning and long-context results are comparatively limited for a premium-priced model.
Benchmarks
Artificial Analysis data refreshed Oct 8, 2026. Arena ratings from the Oct 2, 2026 text leaderboard, with style control. Bars: magenta = reasoning, ink = coding and agents, cyan = instructions and long context.
Artificial Analysis has not published coding or agentic results for xAI: Grok 4.20 yet. Until it does, the model sits below tested models on the Coding and AI Agents lists.
How xAI: Grok 4.20 compares
XAI: Grok 4.20 ranks #74 of 435 AI models we track for overall intelligence. Its 2M-token context window is larger than 100% of the models we list. At $1.25 per million input tokens it is cheaper than 21% of comparable models.
Dark bar = input · light bar = output, scaled to the priciest peer.
1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.
Pricing
| Token Type | Cost per 1M tokens | Cost per 1K tokens |
|---|---|---|
| Input | $1.25 | $0.001250 |
| Output | $2.50 | $0.002500 |
What would xAI: Grok 4.20 cost your business?
Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.
A website chatbot handling around 100 customer conversations a day, a few short messages each.
Full calculator with 697 models → Price Calculator
These numbers get smaller with the right architecture.
We route routine calls to cheap models and save xAI: Grok 4.20 for the hard ones. Most clients cut their estimate by 60-80%.
Talk to our teamAbout xAI: Grok 4.20
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering..
Explore Related Models
Frequently asked questions about xAI: Grok 4.20
How much does xAI: Grok 4.20 cost?
xAI: Grok 4.20 costs $1.25 per million input tokens and $2.50 per million output tokens.
What is the context window of xAI: Grok 4.20?
xAI: Grok 4.20 has a context window of 2,000,000 tokens (2M).
What can xAI: Grok 4.20 do?
xAI: Grok 4.20 supports image/vision input, tool use, and function calling.
Who created xAI: Grok 4.20?
xAI: Grok 4.20 is developed by xAI and was released on March 31, 2026.
Benchmarks from Artificial Analysis, pricing from the OpenRouter API, and verdicts from our own testing.
Last updated: October 8, 2026 8:39 pm