Claude Opus 5 vs Hermes 4 - Llama-3.1 405B: Release Comparison
Comparison of the Claude Opus 5 and Hermes 4 - Llama-3.1 405B releases: 5 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 7 models.
- For intelligence, Claude Opus 5 scores highest: Claude Opus 5 (Adaptive Reasoning, Max Effort) at 51, against Hermes 4 - Llama-3.1 405B (Reasoning) at 7 for Hermes 4 - Llama-3.1 405B.
- For output speed, Claude Opus 5 is fastest: Claude Opus 5 (Adaptive Reasoning, Max Effort) at 50 t/s, against Hermes 4 - Llama-3.1 405B (Non-reasoning) at 43 t/s for Hermes 4 - Llama-3.1 405B.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indices
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Claude Opus 5 (Adaptive Reasoning, Max Effort) | 51 | 55 | 57 | 57 | 53 | 54 | 61 | $5.86 | $3.9 | $5.00 | $25.00 | $0.50 | $7,275 | 73k | 43k | 140M | 50 | 43.51s | 43.51s | 53.51s | 898.49s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) | 50 | 54 | 54 | 56 | 53 | 53 | 60 | $4.88 | $3.9 | $5.00 | $25.00 | $0.50 | $5,868 | 61k | 35k | 111M | 49 | 19.65s | 19.65s | 29.76s | 754.94s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, High Effort) | 48 | 52 | 54 | 55 | 52 | 52 | 58 | $3.61 | $3.9 | $5.00 | $25.00 | $0.50 | $4,332 | 46k | 25k | 81M | 49 | 10.81s | 10.81s | 20.95s | 588.37s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, Medium Effort) | 45 | 49 | 52 | 52 | 50 | 48 | 56 | $2.19 | $3.9 | $5.00 | $25.00 | $0.50 | $2,732 | 29k | 15k | 49M | 48 | 5.62s | 5.62s | 16.09s | 372.96s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Opus 5 (Adaptive Reasoning, Low Effort) | 40 | 44 | 47 | 48 | 45 | 43 | 51 | $1.10 | $3.9 | $5.00 | $25.00 | $0.50 | $1,561 | 15k | 7k | 26M | 46 | 2.41s | 2.41s | 13.30s | 188.57s | - | 1M | Jul 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Hermes 4 - Llama-3.1 405B (Reasoning) | 7 | - | - | - | - | - | - | - | $1.2 | $1.00 | $3.00 | - | - | - | - | - | 42 | 2.46s | 49.60s | 61.38s | - | 406B | 128k | Aug 2025 | - | Yes | Supports: text | Supports: text | |||||
| Hermes 4 - Llama-3.1 405B (Non-reasoning) | 7 | - | - | - | - | - | - | - | $1.2 | $1.00 | $3.00 | - | - | - | - | - | 43 | 2.38s | 2.38s | 13.93s | - | 406B | 128k | Aug 2025 | - | No | Supports: text | Supports: text | |||||