Claude Opus 5 vs Claude 4.5 Haiku: Release Comparison

Comparison of the Claude Opus 5 and Claude 4.5 Haiku releases: 5 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 7 models.

  • For intelligence, Claude Opus 5 scores highest: Claude Opus 5 (Adaptive Reasoning, Max Effort) at 51, against Claude 4.5 Haiku (Reasoning) at 18 for Claude 4.5 Haiku.
  • For output speed, Claude 4.5 Haiku is fastest: Claude 4.5 Haiku (Reasoning) at 88 t/s, against Claude Opus 5 (Adaptive Reasoning, Low Effort) at 54 t/s for Claude Opus 5.
  • For cost per task, Claude 4.5 Haiku is cheapest: Claude 4.5 Haiku (Reasoning) at $0.21, against Claude Opus 5 (Adaptive Reasoning, Low Effort) at $1.10 for Claude Opus 5.

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Opus 5
Pareto line

Cost

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Speed & Latency

Output Speed

Output tokens per second · Higher is better

Capability Scores

Capability Indices

Measures the performance of models on specific capabilities and industries
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench v4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better

Further details

Weights
Provider Benchmarks
Claude Opus 5 (Adaptive Reasoning, Max Effort)
Anthropic logoAnthropic
51
-
1M
$3.9
53
Not available
Amazon BedrockAnthropicGoogle
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Anthropic logoAnthropic
50
-
1M
$3.9
53
Not available
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, High Effort)
Anthropic logoAnthropic
48
-
1M
$3.9
52
Not available
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, Medium Effort)
Anthropic logoAnthropic
45
-
1M
$3.9
52
Not available
AnthropicAmazon BedrockGoogle
Claude Opus 5 (Adaptive Reasoning, Low Effort)
Anthropic logoAnthropic
40
-
1M
$3.9
54
Not available
AnthropicAmazon BedrockGoogle
Claude 4.5 Haiku (Reasoning)
Anthropic logoAnthropic
18
-
200k
$0.8
88
Not available
Amazon BedrockGoogleAnthropicMicrosoft Azure
Claude 4.5 Haiku (Non-reasoning)
Anthropic logoAnthropic
15
-
200k
$0.8
80
Not available
Microsoft AzureAmazon BedrockAnthropicGoogle