Claude Opus 5 vs DeepSeek V4 Pro 0424: Release Comparison

Comparison of the Claude Opus 5 and DeepSeek V4 Pro 0424 releases: 5 and 3 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 8 models.

  • For intelligence, Claude Opus 5 scores highest: Claude Opus 5 (Adaptive Reasoning, Max Effort) at 51, against DeepSeek V4 Pro (Reasoning, Max Effort) at 31 for DeepSeek V4 Pro 0424.
  • For output speed, DeepSeek V4 Pro 0424 is fastest: DeepSeek V4 Pro (Non-reasoning) at 98 t/s, against Claude Opus 5 (Adaptive Reasoning, Max Effort) at 50 t/s for Claude Opus 5.
  • For cost per task, DeepSeek V4 Pro 0424 is cheapest: DeepSeek V4 Pro (Reasoning, Max Effort) at $0.12, against Claude Opus 5 (Adaptive Reasoning, Low Effort) at $1.10 for Claude Opus 5.

Releases compared

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Opus 5
Pareto line

Cost

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Speed & Latency

Output Speed

Output tokens per second · Higher is better

Capability Scores

Capability Indices

Measures the performance of models on specific capabilities and industries
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better

Further details

Input modality
Output modality
Weights
Provider Benchmarks
Claude Opus 5 (Adaptive Reasoning, Max Effort)
Anthropic logoAnthropic
51
55
57
57
53
54
61
$5.86
$3.9
$5.00
$25.00
$0.50
$7,275
73k
43k
140M
50
43.51s
43.51s
53.51s
898.49s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
Amazon BedrockAnthropicGoogle
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Anthropic logoAnthropic
50
54
54
56
53
53
60
$4.88
$3.9
$5.00
$25.00
$0.50
$5,868
61k
35k
111M
49
18.47s
18.47s
28.58s
754.94s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, High Effort)
Anthropic logoAnthropic
48
52
54
55
52
52
58
$3.61
$3.9
$5.00
$25.00
$0.50
$4,332
46k
25k
81M
49
8.58s
8.58s
18.72s
588.37s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, Medium Effort)
Anthropic logoAnthropic
45
49
52
52
50
48
56
$2.19
$3.9
$5.00
$25.00
$0.50
$2,732
29k
15k
49M
48
5.24s
5.24s
15.72s
372.96s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
AnthropicAmazon BedrockGoogle
Claude Opus 5 (Adaptive Reasoning, Low Effort)
Anthropic logoAnthropic
40
44
47
48
45
43
51
$1.10
$3.9
$5.00
$25.00
$0.50
$1,561
15k
7k
26M
46
2.52s
2.52s
13.37s
188.57s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
AnthropicAmazon BedrockGoogle
DeepSeek V4 Pro (Reasoning, Max Effort)
DeepSeek logoDeepSeek
31
35
40
34
33
32
38
$0.12
$0.2
$0.435
$0.87
$0.0036
$277
49k
33k
212M
88
1.64s
51.28s
56.96s
387.69s
1.6T
49B active at inference time
1M
Apr 2026
-
Yes

Supports: text

Supports: text

DeepInfraMicrosoft AzureNovita
+6
DeepSeek V4 Pro (Reasoning, High Effort)
DeepSeek logoDeepSeek
30
-
-
-
-
-
-
-
$0.2
$0.435
$0.87
$0.0036
-
-
-
-
93
1.79s
23.23s
28.62s
-
1.6T
49B active at inference time
1M
Apr 2026
-
Yes

Supports: text

Supports: text

DeepInfraMicrosoft AzureDeepSeek
+5
DeepSeek V4 Pro (Non-reasoning)
DeepSeek logoDeepSeek
21
-
-
-
-
-
-
-
$0.2
$0.435
$0.87
$0.0036
-
-
-
-
98
1.57s
1.57s
6.67s
-
1.6T
49B active at inference time
1M
Apr 2026
-
No

Supports: text

Supports: text

Microsoft AzureBitdeer AINebius
+2