Llama 3.1 Tulu3 405B

Llama 3.1 Tulu3 405B

Allen Institute for AI · Released Jan 30, 2025
Intelligence #616 / 699
14.6 our score
AA Index #310 / 427
8.1 Artificial Analysis
Input Price
Not priced
Output Price
Not priced
Context
Not reported

Analysis Summary

Llama 3.1 Tulu3 405B is a large instruction-tuned open model from the Allen Institute for AI. Within this batch it has the strongest supplied general reasoning, language, and coding results, including meaningful performance on knowledge, code generation, and scientific coding tasks. Its large parameter count may also increase infrastructure demands, although deployment requirements are not provided.

The model is best suited to teams evaluating self-hosted content generation, code assistance, and research workflows that can tolerate human review. It is not supported by tool-use, context, speed, or pricing data, so its fit for autonomous agents and high-volume production remains uncertain. Adopt it where open deployment is central and local evaluation is available.

Assessed August 9, 2026

Editorial notes

Llama 3.1 Tulu3 405B provides the strongest general and coding evidence in this batch, but its measured capability remains well below current frontier systems and operational data is incomplete.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Llama 3.1 Tulu3 405B provides the strongest general and coding evidence in this batch, but its measured capability remains well below current frontier systems and operational data is incomplete.

#616 of 699 overall

Benchmark scores

GPQA Diamond 51.6%
HLE 3.5%
MMLU Pro 71.6%
MATH 500 77.8%
AIME 13.3%
SciCode 30.2%
LiveCodeBench 29.1%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

8.1 Intelligence Index

How Llama 3.1 Tulu3 405B compares

Llama 3.1 Tulu3 405B ranks #310 of 427 AI models we track for overall intelligence. Llama 3.1 Tulu3 405B is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 12% of models #616
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Llama 3.1 Tulu3 405B $0.00 in $0.00 out
Claude Opus 5 $5.00 in $25.00 out
SpaceXAI: Grok 4.6 $2.00 in $6.00 out
Qwen: Qwen3.8 Max $2.00 in $6.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Llama 3.1 Tulu3 405B 0

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence2.1Technical0Value0Content3.8
Performance profile

Strongest on content. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Frequently asked questions about Llama 3.1 Tulu3 405B

How much does Llama 3.1 Tulu3 405B cost?

Llama 3.1 Tulu3 405B is currently available for free via OpenRouter.

Who created Llama 3.1 Tulu3 405B?

Llama 3.1 Tulu3 405B is developed by Allen Institute For AI and was released on January 30, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.