Compare AI Models

Compare AI Models

Select up to 4 models to compare side by side.

Hermes 4 – Llama-3.1 70B (Reasoning) Nous Research
Overview
Provider Nous Research
Our Score 31
Performance Tier Efficient
Best For
Released Aug 2025
Knowledge Cutoff
Dimension Scores
Intelligence 2.1 / 10
Technical 0.0 / 10
Content 4.3 / 10
Value 7.0 / 10
Business Fit 5.0 / 10
Pricing
Input / 1M tokens $0.13
Output / 1M tokens $0.40
Performance
Context Window 0
Throughput (tok/s)
Latency (TTFT ms)
Parameters
Intelligence Indices (Artificial Analysis)
Intelligence 7.9
Coding
Math 68.7
Agentic
Benchmarks
MMLU-Pro 0.8
MMLU
GPQA Diamond 0.7
GPQA (HF)
LiveCodeBench 0.7
HumanEval
MATH-500
AIME '25 0.7
AIME
MATH (HF)
Humanity's Last Exam 0.1
SWE-bench
IFBench 0.3
IFEval

Looking for the full rankings? View the AI Model Leaderboard or try the Price Calculator to estimate monthly costs.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.