Claude Fable 5.1 vs Hermes 4 - Llama-3.1 70B: Release Comparison
Comparison of the Claude Fable 5.1 and Hermes 4 - Llama-3.1 70B releases: 5 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 7 models.
- For intelligence, Claude Fable 5.1 scores highest: Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) at 53, against Hermes 4 - Llama-3.1 70B (Reasoning) at 8 for Hermes 4 - Llama-3.1 70B.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) | 53 | 56 | 60 | 60 | 58 | 56 | 63 | $7.63 | $7.2 | $10.00 | $50.00 | $0.25 | $13,129 | 78k | 47k | 188M | 66 | 298.45s | 298.45s | 306.04s | 742.47s | - | 1M | Sep 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback) | 53 | 56 | 58 | 61 | - | 57 | 62 | $5.98 | $7.2 | $10.00 | $50.00 | $0.25 | $9,063 | 61k | 34k | 121M | 62 | 187.94s | 187.94s | 196.03s | 621.69s | - | 1M | Sep 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback) | 51 | 54 | 56 | 58 | - | 54 | 60 | $3.91 | $7.2 | $10.00 | $50.00 | $0.25 | $5,242 | 38k | 19k | 62M | 56 | 20.56s | 20.56s | 29.52s | 429.41s | - | 1M | Sep 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback) | 49 | 51 | 54 | 57 | - | 51 | 58 | $2.98 | $7.2 | $10.00 | $50.00 | $0.25 | $3,983 | 28k | 12k | 44M | 57 | 8.45s | 8.45s | 17.30s | 310.86s | - | 1M | Sep 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback) | 47 | 49 | 51 | 53 | - | 49 | 55 | $2.37 | $7.2 | $10.00 | $50.00 | $0.25 | $3,158 | 22k | 8k | 33M | 55 | 8.00s | 8.00s | 17.05s | 248.08s | - | 1M | Sep 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Hermes 4 - Llama-3.1 70B (Reasoning) | 8 | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | 70.6B | 128k | Aug 2025 | - | Yes | Supports: text | Supports: text | - | ||||
| Hermes 4 - Llama-3.1 70B (Non-reasoning) | 7 | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | - | 70.6B | 128k | Aug 2025 | - | No | Supports: text | Supports: text | - | ||||