Meta: Llama 4 Maverick

Meta: Llama 4 Maverick

meta-llama · Released Apr 5, 2025
Intelligence #130 / 743
49.5 our score
Speed #182 / 326
81.7 tok/s
Input Price #323 / 743
$0.200 per 1M tokens
Output Price #339 / 743
$0.696 per 1M tokens
Context #42 / 743
1M tokens

Analysis Summary

Llama 4 Maverick is Meta's multimodal model with a 1M token context window, vision, tool use, and function calling. Its measured general reasoning, instruction following, coding, and long-context results are useful for a value-oriented model, while the listed pricing supports broad experimentation and production volume.

It fits large brief analysis, document comparison, visual content processing, codebase context gathering, content transformation, and supervised tool workflows. The agentic measurement is limited, so it should not be trusted with broad autonomous permissions or complex multi-step execution without safeguards. Its context capacity is a major operational asset for consolidating extensive source material, but retrieval quality should still be tested in practice.

Adopt Maverick for affordable multimodal and long-context workloads. It is a strong secondary model for agency operations, while advanced reasoning systems remain preferable for difficult strategy and autonomous engineering.

Assessed August 9, 2026

Editorial notes

Llama 4 Maverick offers a 1M token context, vision, function calling, and low pricing, with useful reasoning and long-context measurements. Its agentic capability is limited, so it suits supervised document and content workflows more than autonomous agents.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Llama 4 Maverick offers a 1M token context, vision, function calling, and low pricing, with useful reasoning and long-context measurements. Its agentic capability is limited, so it suits supervised document and content workflows more than autonomous agents.

#130 of 743 overall Down 2 this week

Benchmark scores

GPQA Diamond 67.1%
HLE 4.8%
MMLU Pro 80.9%
MATH 500 88.9%
AIME 39%
AIME 2025 19.3%
SciCode 33.1%
LiveCodeBench 39.7%
TerminalBench Hard 6.8%
τ²-Bench 17.8%
IFBench 43%
LCR 46%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

14.5 Intelligence Index·16.3 Coding Index·1.2 Agentic Index·19.3 Math Index

How Meta: Llama 4 Maverick compares

Meta: Llama 4 Maverick ranks #225 of 432 AI models we track for overall intelligence, #154 of 205 for coding, #174 of 187 for agentic tasks. Its 1M-token context window is larger than 94% of the models we list. At $0.20 per million input tokens it is cheaper than 57% of comparable models.

Position in the field
Intelligence: smarter than 83% of models #130
Speed: faster than 44% of models #182
Price: cheaper than 57% of models #323
Context: larger than 94% of models #42
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Meta: Llama 4 Maverick $0.20 in $0.70 out
OpenAI: GPT-5.6 Sol $2.00 in $10.00 out
Claude Opus 5 $5.00 in $25.00 out
SpaceXAI: Grok 4.6 $2.00 in $6.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Meta: Llama 4 Maverick 1M

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence3Technical1.8Value8.3Content6
Performance profile

Strongest on value. The pulled-in technical corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $0.20 $0.000200
Output $0.70 $0.000696

What would Meta: Llama 4 Maverick cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$0/mo Meta: Llama 4 Maverick

Full calculator with 743 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Meta: Llama 4 Maverick for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

About Meta: Llama 4 Maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward..

Frequently asked questions about Meta: Llama 4 Maverick

How much does Meta: Llama 4 Maverick cost?

Meta: Llama 4 Maverick costs $0.20 per million input tokens and $0.70 per million output tokens.

What is the context window of Meta: Llama 4 Maverick?

Meta: Llama 4 Maverick has a context window of 1,048,576 tokens (1M).

Is Meta: Llama 4 Maverick good for coding?

On our coding benchmark index, Meta: Llama 4 Maverick ranks #154 of 205 models, placing it in the broader range of the field for code generation and debugging.

What can Meta: Llama 4 Maverick do?

Meta: Llama 4 Maverick supports image/vision input, tool use, and function calling.

Who created Meta: Llama 4 Maverick?

Meta: Llama 4 Maverick is developed by Meta and was released on April 5, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.