Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)

Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)

NVIDIA · Released May 20, 2025
Intelligence #687 / 756
11.1 our score
AA Index #328 / 435
2.9 Artificial Analysis
Input Price
Not priced
Output Price
Not priced
Context
Not reported

Analysis Summary

Llama 3.1 Nemotron Nano 4B Reasoning is a compact NVIDIA model with a low general intelligence result and limited measured capability across coding, mathematics, instruction following, and tool use. The small model profile may be useful where lightweight deployment is important, but no pricing or context information is provided.

For an agency, suitable applications are narrow classification, simple extraction, and constrained experimentation rather than open-ended content generation or autonomous agents. Its instruction-following and tool-use results are too limited for dependable client-facing automation, while the coding profile does not support serious software engineering. Further testing would be needed to establish whether deployment efficiency compensates for its capability limits.

Assessed September 7, 2026

Editorial notes

Llama 3.1 Nemotron Nano 4B Reasoning is a small NVIDIA model with limited general reasoning and modest coding-related results. Its low instruction-following and tool-use measurements restrict it to narrow experiments rather than client-critical workflows.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Llama 3.1 Nemotron Nano 4B Reasoning is a small NVIDIA model with limited general reasoning and modest coding-related results. Its low instruction-following and tool-use measurements restrict it to narrow experiments rather than client-critical workflows.

#687 of 756 overall Down 56 this week

Benchmark scores

GPQA Diamond 40.8%
HLE 5.1%
MMLU Pro 55.6%
MATH 500 94.7%
AIME 70.7%
AIME 2025 50%
SciCode 10.1%
LiveCodeBench 49.3%
τ²-Bench 11.7%
IFBench 25.5%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

2.9 Intelligence Index·50 Math Index

How Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) compares

Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) ranks #328 of 435 AI models we track for overall intelligence. Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) is currently free to use via OpenRouter.

Position in the field
Intelligence: smarter than 9% of models #687
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) $0.00 in $0.00 out
OpenAI: GPT-6 Astra $10.00 in $50.00 out
Anthropic: Claude Fable 5.1 $10.00 in $50.00 out
Claude Opus 5 $5.00 in $25.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) 0

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence1.3Technical0Value0Content3
Performance profile

Strongest on business fit. The pulled-in value corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Frequently asked questions about Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)

How much does Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) cost?

Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) is currently available for free via OpenRouter.

Who created Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)?

Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning) is developed by NVIDIA and was released on May 20, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.