Claude Sonnet 5 vs gpt-oss-20b: Release Comparison

Comparison of the Claude Sonnet 5 and gpt-oss-20b releases: 6 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 8 models.

  • For intelligence, Claude Sonnet 5 scores highest: Claude Sonnet 5 (Adaptive Reasoning, Max Effort) at 38, against gpt-oss-20b (low) at 10 for gpt-oss-20b.
  • For output speed, gpt-oss-20b is fastest: gpt-oss-20b (low) at 266 t/s, against Claude Sonnet 5 (Adaptive Reasoning, Max Effort) at 84 t/s for Claude Sonnet 5.
  • For cost per task, gpt-oss-20b is cheapest: gpt-oss-20b (high) at $0.01, against Claude Sonnet 5 (Adaptive Reasoning, Low Effort) at $0.51 for Claude Sonnet 5.

Releases compared

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Sonnet 5
Pareto line

Cost

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Speed & Latency

Output Speed

Output tokens per second · Higher is better

Capability Scores

Capability Indices

Measures the performance of models on specific capabilities and industries
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better

Further details

Input modality
Output modality
Weights
Provider Benchmarks
Claude Sonnet 5 (Adaptive Reasoning, Max Effort)
Anthropic logoAnthropic
38
41
40
43
43
40
50
$5.09
$1.5
$2.00
$10.00
$0.20
$6,998
118k
88k
367M
84
195.83s
195.83s
201.79s
877.37s
-
1M
Jun 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
GoogleMicrosoft AzureAnthropicAmazon Bedrock
Claude Sonnet 5 (Adaptive Reasoning, Xhigh Effort)
Anthropic logoAnthropic
35
38
37
39
-
36
46
$2.87
$1.5
$2.00
$10.00
$0.20
$3,255
65k
44k
135M
62
22.30s
22.30s
30.37s
672.64s
-
1M
Jun 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
Anthropic
Claude Sonnet 5 (Adaptive Reasoning, High Effort)
Anthropic logoAnthropic
32
34
34
35
-
32
42
$1.79
$1.5
$2.00
$10.00
$0.20
$2,086
44k
28k
89M
58
2.30s
2.30s
10.90s
460.65s
-
1M
Jun 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
Anthropic
Claude Sonnet 5 (Non-reasoning, High Effort)
Anthropic logoAnthropic
29
-
-
-
-
-
-
-
$1.5
$2.00
$10.00
$0.20
-
-
-
-
58
1.08s
1.08s
9.72s
-
-
1M
Jun 2026
-
No

Supports: text and image

Supports: text

Not available
-
AnthropicAmazon Bedrock
Claude Sonnet 5 (Adaptive Reasoning, Medium Effort)
Anthropic logoAnthropic
28
31
31
32
-
29
38
$1.00
$1.5
$2.00
$10.00
$0.20
$1,218
27k
15k
51M
57
1.50s
1.50s
10.34s
285.97s
-
1M
Jun 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
Anthropic
Claude Sonnet 5 (Adaptive Reasoning, Low Effort)
Anthropic logoAnthropic
25
27
27
29
-
26
34
$0.51
$1.5
$2.00
$10.00
$0.20
$653
15k
7k
25M
56
1.21s
1.21s
10.17s
161.65s
-
1M
Jun 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
Anthropic
gpt-oss-20b (low)
OpenAI logoOpenAI
10
-
-
-
-
-
-
-
$0.1
$0.07
$0.20
-
-
-
-
-
266
0.85s
8.38s
10.26s
-
21B
3.6B active at inference time
131k
Aug 2025
May 2025
Yes

Supports: text

Supports: text

GroqLightning AIHyperbolic
+5
gpt-oss-20b (high)
OpenAI logoOpenAI
9
7
5
7
7
12
11
$0.01
$0.1
$0.06
$0.19
-
$29
18k
16k
70M
222
0.76s
9.79s
12.04s
76.54s
21B
3.6B active at inference time
131k
Aug 2025
May 2024
Yes

Supports: text

Supports: text

CoreWeaveAmazon BedrockDeepInfra
+6