Release Comparison Tool
Compare up to five model releases side by side, every variant of each on intelligence, pricing, output speed, latency, context window and more.
For details about how we measure and evaluate models, see our Methodology page.
| Claude Opus 5.5 (max with fallback) | Claude Opus 5.5 (xhigh with fallback) | Claude Opus 5.5 (high with fallback) | GPT-6 Astra (max) | GPT-6 Astra (xhigh) | Claude Opus 5.5 (medium with fallback) | GPT-6 Astra (high) | GPT-6 Astra (medium) | GPT-6 Astra (low) | Claude Opus 5.5 (low with fallback) | |
|---|---|---|---|---|---|---|---|---|---|---|
| Intelligence Index | 58 | 56 | 54 | 53 | 52 | 51 | 51 | 50 | 46 | 42 |
| Finance & Accounting | 61 | 58 | 56 | 55 | 54 | 54 | 53 | 52 | 49 | 46 |
| Strategy & Ops | 64 | 62 | 59 | 57 | 57 | 57 | 55 | 54 | 50 | 48 |
| Legal | 63 | 61 | 59 | 59 | 58 | 57 | 58 | 57 | 53 | 52 |
| Healthcare & Medical | 61 | -- | -- | 52 | -- | -- | -- | -- | -- | -- |
| Cost per TaskUSD | $5.98 | $3.46 | $1.82 | $3.26 | $2.31 | $1.34 | $1.73 | $1.54 | $0.82 | $0.55 |
| Output SpeedTokens/s | 92 | 79 | 72 | 59 | 55 | 73 | 53 | 51 | 51 | 71 |
| Time to First Tokens | 710.36 | 147.01 | 34.40 | 276.50 | 142.98 | 26.57 | 67.00 | 4.50 | 2.65 | 4.50 |
| Context Window | 1M | 1M | 1M | 1M | 1M | 1M | 1M | 1M | 1M | 1M |
| Creator | ||||||||||
| License | Proprietary | Proprietary | Proprietary | Proprietary | Proprietary | Proprietary | Proprietary | Proprietary | Proprietary | Proprietary |
| Input Modality | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image | Supports: text and image |
| Output Modality | Supports: text | Supports: text | Supports: text | Supports: text | Supports: text | Supports: text | Supports: text | Supports: text | Supports: text | Supports: text |
| Providers | ||||||||||
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better