Claude Opus 5 vs Hermes 4 - Llama-3.1 405B: Release Comparison
Comparison of the Claude Opus 5 and Hermes 4 - Llama-3.1 405B releases: 5 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 7 models.
- For intelligence, Claude Opus 5 scores highest: Claude Opus 5 (Adaptive Reasoning, Max Effort) at 51, against Hermes 4 - Llama-3.1 405B (Reasoning) at 7 for Hermes 4 - Llama-3.1 405B.
- For output speed, Claude Opus 5 is fastest: Claude Opus 5 (Adaptive Reasoning, Medium Effort) at 57 t/s, against Hermes 4 - Llama-3.1 405B (Non-reasoning) at 41 t/s for Hermes 4 - Llama-3.1 405B.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Claude Opus 5 (Adaptive Reasoning, Max Effort) | 51 | 55 | 57 | 57 | 53 | 54 | 61 | $5.86 | $3.9 | $5.00 | $25.00 | $0.50 | $7,275 | 73k | 43k | 140M | 55 | 62.61s | 62.61s | 71.64s | 803.97s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) | 50 | 54 | 54 | 56 | 53 | 53 | 60 | $4.88 | $3.9 | $5.00 | $25.00 | $0.50 | $5,868 | 61k | 35k | 111M | 54 | 26.58s | 26.58s | 35.86s | 700.83s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, High Effort) | 48 | 51 | 53 | 54 | 52 | 52 | 58 | $3.61 | $3.9 | $5.00 | $25.00 | $0.50 | $4,332 | 46k | 25k | 81M | 56 | 16.08s | 16.08s | 25.04s | 508.17s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, Medium Effort) | 45 | 49 | 51 | 52 | 50 | 47 | 56 | $2.19 | $3.9 | $5.00 | $25.00 | $0.50 | $2,732 | 29k | 15k | 49M | 57 | 5.20s | 5.20s | 13.99s | 308.27s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, Low Effort) | 39 | 44 | 46 | 48 | 45 | 42 | 50 | $1.10 | $3.9 | $5.00 | $25.00 | $0.50 | $1,561 | 15k | 7k | 26M | 57 | 3.04s | 3.04s | 11.84s | 156.93s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Hermes 4 - Llama-3.1 405B (Reasoning) | 7 | - | - | - | - | - | - | - | $1.2 | $1.00 | $3.00 | - | - | - | - | - | 40 | 2.46s | 52.39s | 64.88s | - | 406B | 128k | Aug 2025 | - | Yes | Supports: text | Supports: text | |||||
| Hermes 4 - Llama-3.1 405B (Non-reasoning) | 7 | - | - | - | - | - | - | - | $1.2 | $1.00 | $3.00 | - | - | - | - | - | 41 | 2.36s | 2.36s | 14.47s | - | 406B | 128k | Aug 2025 | - | No | Supports: text | Supports: text | |||||