Hermes 4 – Llama-3.1 405B (Reasoning)

Hermes 4 – Llama-3.1 405B (Reasoning)

Nous Research · Released Aug 27, 2025
Intelligence #258 / 650
35.4 our score
Speed #281 / 296
30.1 tok/s
Input Price #495 / 688
$1.00 per 1M tokens
Output Price #484 / 688
$3.00 per 1M tokens
Context
Not reported

Analysis Summary

This reasoning-enabled Hermes 4 variant from Nous Research posts strong math and coding benchmark results, suggesting real strength in structured problem-solving despite a lower general intelligence measure.

It could suit technical teams needing math-heavy or coding support on Llama-based infrastructure, though pricing is high for the overall capability on offer, and general content quality lags behind current flagships.

Assessed July 29, 2026

Editorial notes

Hermes 4 405B Reasoning shows strong math and coding benchmark results despite a low overall intelligence index, at a premium price point.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Hermes 4 405B Reasoning shows strong math and coding benchmark results despite a low overall intelligence index, at a premium price point.

#258 of 650 overall

Benchmark scores

GPQA Diamond 72.7%
HLE 10.3%
MMLU Pro 82.9%
AIME 2025 69.7%
SciCode 25.2%
LiveCodeBench 68.6%
TerminalBench Hard 11.4%
τ²-Bench 22.2%
IFBench 32.7%
LCR 20.7%

Magenta = intelligence · Ink = technical/agentic · Cyan = content & long-context · Grey = community benchmarks. Data: Artificial Analysis, Hugging Face.

9 Intelligence Index·69.7 Math Index

How Hermes 4 – Llama-3.1 405B (Reasoning) compares

Hermes 4 – Llama-3.1 405B (Reasoning) ranks #263 of 401 AI models we track for overall intelligence. At $1.00 per million input tokens it is cheaper than 28% of comparable models.

Position in the field
Intelligence: smarter than 60% of models #258
Speed: faster than 5% of models #281
Price: cheaper than 28% of models #495
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Hermes 4 – Llama-3.1 405B (Reasoning) $1.00 in $3.00 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out
Claude Opus 5 $5.00 in $25.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Hermes 4 – Llama-3.1 405B (Reasoning) 0

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence2.6Technical2.4Value6Content3.5
Performance profile

Strongest on value. The pulled-in technical corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $1.00 $0.001000
Output $3.00 $0.003000

What would Hermes 4 – Llama-3.1 405B (Reasoning) cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$0/mo Hermes 4 – Llama-3.1 405B (Reasoning)

Full calculator with 688 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Hermes 4 – Llama-3.1 405B (Reasoning) for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

Frequently asked questions about Hermes 4 – Llama-3.1 405B (Reasoning)

How much does Hermes 4 – Llama-3.1 405B (Reasoning) cost?

Hermes 4 – Llama-3.1 405B (Reasoning) costs $1.00 per million input tokens and $3.00 per million output tokens.

Who created Hermes 4 – Llama-3.1 405B (Reasoning)?

Hermes 4 – Llama-3.1 405B (Reasoning) is developed by Nous Research and was released on August 27, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.