LLM Leaderboard - Comparison of AI models from OpenAI, Anthropic, Google, SpaceXAI & others

Comparison and ranking the performance of over 250 AI models (LLMs) across key metrics including intelligence, price, performance and speed (output speed - tokens per second & latency - TTFT), context window & others.

For more details including relating to our methodology, see our FAQs.

Intelligence

Claude Opus 5 (max) and Claude Opus 5 (xhigh) are the highest intelligence models, followed by Claude Fable 5 (with fallback) and Claude Opus 5 (high).

Output Speed

Celeris-1 and Mercury 2 are the fastest models, followed by Ling 3.0 Flash and Step 3.7 Flash.

Latency

Gemini 2.5 Flash-Lite and Command A+ are the lowest latency models, followed by Gemini 2.5 Flash and NVIDIA Nemotron 3 Nano.

Cost per Task

GPT-5.6 Luna (low) and MiMo-V2.5 have the lowest cost per task, followed by Llama 4 Scout and GPT-5.6 Luna (medium).

Context Window

Llama 4 Scout and Grok 4.20 0309 support the largest context windows, followed by Gemini 1.5 Pro (May) and Grok 4.1 Fast.

Further Analysis
Claude Opus 5 (max)
1M
AnthropicAnthropic
63
$2.34
57
33.52
42.28
Claude Opus 5 (xhigh)
1M
AnthropicAnthropic
63
$1.80
58
23.72
32.28
Claude Fable 5 (with fallback)
1M
AnthropicAnthropic
62
$3.14
69
137.13
144.38
Claude Opus 5 (high)
1M
AnthropicAnthropic
61
$1.23
58
12.73
21.37
GPT-5.6 Sol (max)
1M
OpenAIOpenAI
61
$1.23
69
133.61
140.84
Kimi K3 (max)
1.05M
KimiKimi
60
$0.84
44
2.42
59.21
GPT-5.6 Sol (xhigh)
1M
OpenAIOpenAI
59
$0.81
69
49.89
57.15
Claude Opus 5 (medium)
1M
AnthropicAnthropic
59
$0.72
57
5.67
14.41
Qwen3.8 Max
1M
AlibabaAlibaba
58
$1.13
82
2.76
33.41
GPT-5.6 Sol (high)
1M
OpenAIOpenAI
57
$0.55
67
15.76
23.20
Muse Spark 1.2 (xhigh)
1.05M
MetaMeta
57
$0.40
--
--
--
GPT-5.6 Terra (max)
1M
OpenAIOpenAI
57
$0.51
151
167.74
171.06
Grok 4.5 (high)
500k
SpaceXAISpaceXAI
56
$0.36
60
10.55
18.82
GPT-5.6 Sol (medium)
1M
OpenAIOpenAI
56
$0.37
68
4.39
11.75
Claude Sonnet 5 (max)
1M
AnthropicAnthropic
55
$1.72
82
152.45
158.57
GPT-5.6 Terra (xhigh)
1M
OpenAIOpenAI
53
$0.31
137
21.61
25.24
GLM-5.2 (max)
1M
Z AIZ AI
53
$0.30
143
1.40
18.83
Claude Opus 5 (low)
1M
AnthropicAnthropic
52
$0.43
57
3.42
12.22
GPT-5.6 Luna (max)
1M
OpenAIOpenAI
52
$0.05
202
143.87
146.36
DeepSeek V4 Flash 0731 (max)
1M
DeepSeekDeepSeek
52
$0.03
141
1.39
19.08
Gemini 3.6 Flash
1M
GoogleGoogle
52
$0.56
238
24.13
26.23
GPT-5.6 Sol (low)
1M
OpenAIOpenAI
51
$0.23
65
3.48
11.12
GPT-5.6 Terra (high)
1M
OpenAIOpenAI
50
$0.22
138
2.65
6.27
GPT-5.6 Luna (xhigh)
1M
OpenAIOpenAI
50
$0.03
199
39.34
41.86
Kimi K3 (low)
1.05M
KimiKimi
48
$0.24
44
2.82
59.66
Gemini 3.1 Pro Preview
1M
GoogleGoogle
48
$0.33
136
44.36
48.03
GPT-5.6 Luna (high)
1M
OpenAIOpenAI
47
$0.02
189
8.42
11.06
GPT-5.6 Terra (medium)
1M
OpenAIOpenAI
47
$0.12
133
1.60
5.36
Gemini 3.5 Flash (medium)
1M
GoogleGoogle
47*
--
188
25.99
28.64
GPT-5.3 Codex (xhigh)
400k
OpenAIOpenAI
46*
--
145
71.97
75.43
MiniMax-M3
1M
MiniMaxMiniMax
45
$0.14
96
1.34
27.33
DeepSeek V4 Pro (max)
1M
DeepSeekDeepSeek
45
$0.05
79
1.89
63.76
Motif 3 (Beta)
262k
Motif TechnologiesMotif Technologies
45
--
--
--
--
DeepSeek V4 Pro (high)
1M
DeepSeekDeepSeek
44
$0.04
79
1.73
33.18
Kimi K2.7 Code
256k
KimiKimi
43
$0.22
43
2.84
66.00
MiMo-V2.5-Pro
1M
XiaomiXiaomi
43
$0.03
77
2.87
35.32
Claude Sonnet 5 (Non-reasoning)
1M
AnthropicAnthropic
43
$0.42
73
1.81
8.67
Inkling
1M
Thinking MachinesThinking Machines
42
$0.34
79
1.86
33.68
Hy3
256k
TencentTencent
42
$0.04
66
2.81
40.79
GPT-5.6 Sol (Non-reasoning)
1M
OpenAIOpenAI
42
$0.24
64
1.18
8.99
Nex-N2-Pro
262k
Nex AGINex AGI
42
--
132
1.66
20.55
GPT-5.6 Terra (low)
1M
OpenAIOpenAI
41
$0.09
136
1.50
5.17
Inkling Small
1M
Thinking MachinesThinking Machines
41
$0.07
144
1.56
18.95
JT-4.1 Flash 236B A21B
256k
China MobileChina Mobile
40
--
--
--
--
Agnes 2.5 Pro Alpha
1M
Sapiens AISapiens AI
40
--
132
2.52
21.43
Qwen3.7 Plus
1M
AlibabaAlibaba
39
$0.24
56
2.28
46.55
GPT-5.6 Luna (medium)
1M
OpenAIOpenAI
39
$0.01
181
2.25
5.02
Nemotron 3 Ultra
262k
NVIDIANVIDIA
38
$0.38
171
2.04
18.23
MiMo-V2.5
1M
XiaomiXiaomi
38
$0.01
87
2.85
31.61
Ling 3.0 Flash
262k
InclusionAIInclusionAI
38
$0.04
415
2.09
8.12
Qwen3.6 27B
262k
AlibabaAlibaba
38
$0.29
60
3.84
107.05
Gemini 3.5 Flash-Lite
1M
GoogleGoogle
37
$0.10
397
9.84
11.10
MiMo-V2-Omni-0327
256k
XiaomiXiaomi
37*
--
--
--
--
Grok 4.3 (medium)
1M
SpaceXAISpaceXAI
37*
--
139
13.73
17.34
Grok 4.3 (low)
1M
SpaceXAISpaceXAI
36*
--
131
7.46
11.28
MiMo-V2-Omni
256k
XiaomiXiaomi
36*
--
--
--
--
Gemini 3.5 Flash (minimal)
1M
GoogleGoogle
36*
--
167
0.91
3.90
Claude Sonnet 4.6 (Non-reasoning, Low Effort)
1M
AnthropicAnthropic
35*
--
44
1.49
12.96
GLM-5.2
1M
Z AIZ AI
35
--
123
1.90
5.98
GPT-5.6 Terra (Non-reasoning)
1M
OpenAIOpenAI
35
$0.10
131
0.76
4.57
Qwen3.5 397B A17B
262k
AlibabaAlibaba
34
$0.36
77
2.36
50.40
MiMo-V2-Flash (Feb 2026)
256k
XiaomiXiaomi
34*
--
--
--
--
LongCat 2.0
1M
LongCatLongCat
34
$0.12
44
2.78
60.23
KAT-Coder-Pro V2
256k
KwaiKATKwaiKAT
34
--
105
1.66
6.40
GPT-5.6 Luna (low)
1M
OpenAIOpenAI
34
$0.01
176
2.21
5.05
Qwen3.5 122B A10B
262k
AlibabaAlibaba
33
$0.25
135
2.36
20.83
Qwen3.5 397B A17B
262k
AlibabaAlibaba
33*
--
78
2.26
8.66
Qwen3.6 35B A3B
262k
AlibabaAlibaba
32
$0.19
149
2.26
41.71
DeepSeek V4 Pro
1M
DeepSeekDeepSeek
32*
--
75
1.68
8.33
G9v3-39A5B
131k
AI9StarsAI9Stars
31
$0.00
--
--
--
Qwen3.5 Omni Plus
256k
AlibabaAlibaba
31*
--
54
2.35
11.63
Qwen3.6 27B
262k
AlibabaAlibaba
31
$0.40
61
3.89
12.10
Ring-2.6-1T
262k
InclusionAIInclusionAI
31
$0.37
135
3.32
21.78
o3
200k
OpenAIOpenAI
31*
--
171
6.51
9.43
Step 3.7 Flash
262k
StepFunStepFun
31
$0.09
408
0.84
6.96
Mistral Medium 3.5
256k
MistralMistral
30
$0.46
128
2.55
22.14
Claude 4.5 Haiku
200k
AnthropicAnthropic
30
$0.22
101
16.97
21.93
Gemma 4 31B
256k
GoogleGoogle
30
$0.00
34
1.13
66.15
DeepSeek V4 Flash
1M
DeepSeekDeepSeek
29*
--
--
--
--
GPT-5.5 Instant (June 2026)
400k
OpenAIOpenAI
29
$0.54
162
1.21
16.63
JT-35B-Flash
256k
China MobileChina Mobile
29*
--
--
--
--
MiMo-V2.5-Pro
1M
XiaomiXiaomi
28*
--
73
3.10
9.95
Qwen3.5 122B A10B
262k
AlibabaAlibaba
28
$0.20
152
2.32
5.61
GPT-5.6 Luna (Non-reasoning)
1M
OpenAIOpenAI
27
$0.01
180
0.71
3.49
Doubao Seed Code
256k
ByteDance SeedByteDance Seed
26*
--
--
--
--
Gemma 4 26B A4B
256k
GoogleGoogle
26
$0.04
--
--
--
Nemotron 3 Super
1M
NVIDIANVIDIA
26
$0.23
137
1.96
20.15
Grok 4.3 (Non-reasoning)
1M
SpaceXAISpaceXAI
25
$0.29
122
0.71
4.81
MiMo-V2-Flash
256k
XiaomiXiaomi
25
--
--
--
--
Qwen3.6 35B A3B
262k
AlibabaAlibaba
25
$0.63
178
2.45
5.25
Ling 3.0 Tiny
262k
InclusionAIInclusionAI
25
$0.00
169
2.48
17.28
Qwen3.5 35B A3B
262k
AlibabaAlibaba
24
$0.24
172
2.10
5.00
Claude 4.5 Haiku
200k
AnthropicAnthropic
24*
--
90
0.90
6.48
gpt-oss-120b (high)
131k
OpenAIOpenAI
24
$0.07
176
0.88
15.11
Command A+
192k
CohereCohere
23
$0.00
182
0.39
14.10
K-EXAONE
256k
LG AI ResearchLG AI Research
22
--
--
--
--
ERNIE 5.0 Thinking Preview
128k
BaiduBaidu
22*
--
--
--
--
Gemma 4 12B
256k
GoogleGoogle
22
$0.08
111
2.39
24.92
Gemma 4 31B
256k
GoogleGoogle
22
$0.04
60
2.16
10.43
Nova 2.0 Pro Preview (medium)
256k
AmazonAmazon
22
$0.18
122
16.14
36.58
Mercury 2
128k
InceptionInception
22
$0.08
1,061
4.45
4.93
Qwen3.5 9B
262k
AlibabaAlibaba
22
$0.24
91
1.72
29.18
Qwen3 Coder Next
256k
AlibabaAlibaba
21
$0.34
154
1.27
4.51
Nova 2.0 Omni (medium)
1M
AmazonAmazon
21*
--
--
--
--
Apriel-v1.6-15B-Thinker
128k
ServiceNowServiceNow
21*
--
--
--
--
Qwen3.5 9B
262k
AlibabaAlibaba
21*
--
94
0.84
6.16
EXAONE 4.5 33B
262k
LG AI ResearchLG AI Research
21
--
--
--
--
Gemma 4 26B A4B
256k
GoogleGoogle
20*
--
60
1.08
9.44
Qwen3.5 4B
262k
AlibabaAlibaba
20*
--
36
0.93
70.65
North Mini Code
256k
CohereCohere
20
$0.00
27
0.57
92.38
Nova 2.0 Pro Preview (low)
256k
AmazonAmazon
20
$0.23
128
12.01
31.50
Mistral Small 4
256k
MistralMistral
20
$0.10
166
0.82
15.86
Devstral 2
256k
MistralMistral
19
$0.00
50
1.46
11.54
Nova 2.0 Lite (medium)
1M
AmazonAmazon
19*
--
151
17.62
34.16
Qwen3.5 Omni Flash
256k
AlibabaAlibaba
19*
--
275
1.84
3.66
JT-MINI
128k
China MobileChina Mobile
19*
--
--
--
--
Trinity Large Thinking
512k
Arcee AIArcee AI
19
$0.16
205
1.08
13.30
Nova 2.0 Lite (high)
1M
AmazonAmazon
18
$0.24
140
20.34
38.23
HyperNova 60B 2605
131k
Multiverse ComputingMultiverse Computing
18
$0.02
360
0.96
7.91
Magistral Medium 1.2
128k
MistralMistral
18
$0.89
135
2.24
20.74
Nova 2.0 Lite (low)
1M
AmazonAmazon
18*
--
153
9.40
25.73
Nemotron Cascade 2 30B A3B
1M
NVIDIANVIDIA
18
--
--
--
--
Devstral Small 2
256k
MistralMistral
18
$0.00
130
2.34
6.20
LongCat Flash Lite
256k
LongCatLongCat
17*
--
--
--
--
K2 Think V2
262k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
17
--
--
--
--
HyperCLOVA X SEED Think (32B)
128k
NaverNaver
17*
--
--
--
--
K-EXAONE
256k
LG AI ResearchLG AI Research
17*
--
--
--
--
Qwen3 Next 80B A3B
262k
AlibabaAlibaba
17
$0.19
206
2.29
14.42
Nova 2.0 Omni (low)
1M
AmazonAmazon
17*
--
--
--
--
Mi:dm K 2.5 Pro
128k
Korea TelecomKorea Telecom
17*
--
--
--
--
G9v3-3B
131k
AI9StarsAI9Stars
16
$0.00
--
--
--
Qwen3.5 4B
262k
AlibabaAlibaba
16*
--
31
0.72
16.66
Mistral Large 3
256k
MistralMistral
16
$0.08
43
1.27
12.89
INTELLECT-3
131k
Prime IntellectPrime Intellect
16*
--
--
--
--
Solar Open 100B
128k
UpstageUpstage
15*
--
--
--
--
gpt-oss-20b (high)
131k
OpenAIOpenAI
15
$0.02
123
0.99
21.24
Nemotron 3 Nano Omni 30B A3B Reasoning
256k
NVIDIANVIDIA
15*
--
318
1.01
8.88
gpt-oss-120b (low)
131k
OpenAIOpenAI
15
$0.02
240
0.90
11.32
NVIDIA Nemotron 3 Nano
1M
NVIDIANVIDIA
15
$0.02
230
1.24
12.10
Solar Pro 3
128k
UpstageUpstage
14
$0.15
144
2.33
19.71
Llama 4 Maverick
1M
MetaMeta
14
$0.04
137
0.99
4.63
Nova 2.0 Pro Preview
256k
AmazonAmazon
14
$0.26
108
1.04
5.68
gpt-oss-20b (low)
131k
OpenAIOpenAI
14*
--
125
1.05
20.98
K2-V2 (high)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
14*
--
--
--
--
DiffusionGemma 26B A4B
256k
GoogleGoogle
14
--
--
--
--
Qwen3 Next 80B A3B
262k
AlibabaAlibaba
14*
--
200
2.17
4.67
Gemma 4 12B (Non-reasoning)
262k
GoogleGoogle
13*
--
111
2.52
7.01
Motif-2-12.7B
128k
Motif TechnologiesMotif Technologies
13*
--
--
--
--
Nova Premier
1M
AmazonAmazon
13*
--
32
2.85
18.65
Celeris-1
131k
CelerisCeleris
12
$0.26
1,549
0.63
0.95
K2-V2 (medium)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
12*
--
--
--
--
Llama Nemotron Super 49B v1.5
128k
NVIDIANVIDIA
12*
--
144
3.30
20.69
Gemma 4 E4B
128k
GoogleGoogle
12
$0.01
76
0.87
33.90
Mistral Small 4
256k
MistralMistral
12*
--
152
0.77
4.06
Tri-21B-Think
32k
Trillion LabsTrillion Labs
12*
--
--
--
--
MiniCPM5-1B
128k
OpenBMBOpenBMB
12*
--
--
--
--
Sarvam 105B (high)
128k
SarvamSarvam
12*
--
--
--
--
Nova 2.0 Lite
1M
AmazonAmazon
12*
--
170
1.06
4.00
MiniCPM5-1B
128k
OpenBMBOpenBMB
12*
--
--
--
--
Magistral Small 1.2
128k
MistralMistral
11
$0.30
146
0.81
17.91
Ministral 3 14B
256k
MistralMistral
11
$0.16
98
0.88
5.98
Nanbeige4.1-3B
256k
NanbeigeNanbeige
11
--
--
--
--
EXAONE 4.0 32B
131k
LG AI ResearchLG AI Research
11*
--
--
--
--
Nova 2.0 Omni
1M
AmazonAmazon
10*
--
--
--
--
Llama 4 Scout
10M
MetaMeta
10
$0.01
106
0.77
5.50
Hermes 4 70B
128k
Nous ResearchNous Research
10*
--
92
1.36
28.50
Gemma 4 E2B
128k
GoogleGoogle
10
--
--
--
--
Falcon-H1R-7B
256k
TII UAETII UAE
10*
--
--
--
--
Qwen3 Omni 30B A3B
65.5k
AlibabaAlibaba
10*
--
94
1.96
28.68
Step3 VL 10B
65.5k
StepFunStepFun
9*
--
--
--
--
Llama 3.3 70B
128k
MetaMeta
9
$0.10
82
1.77
7.87
Ministral 3 8B
256k
MistralMistral
9
$0.18
126
0.70
4.67
Granite 4.1 30B
131k
IBMIBM
9
--
--
--
--
Llama Nemotron Ultra
128k
NVIDIANVIDIA
9*
--
52
2.37
50.10
ERNIE 4.5 300B A47B
131k
BaiduBaidu
9*
--
--
--
--
Hermes 4 405B
128k
Nous ResearchNous Research
9*
--
41
2.36
63.69
NVIDIA Nemotron 3 Nano 4B
262k
NVIDIANVIDIA
9
--
--
--
--
NVIDIA Nemotron Nano 12B v2 VL
128k
NVIDIANVIDIA
9*
--
113
4.85
27.05
Gemma 4 E4B
128k
GoogleGoogle
9*
--
77
0.75
7.25
NVIDIA Nemotron Nano 9B V2
131k
NVIDIANVIDIA
9*
--
139
4.87
22.84
Hermes 4 405B
128k
Nous ResearchNous Research
9*
--
41
2.32
14.58
Llama Nemotron Super 49B v1.5
128k
NVIDIANVIDIA
9*
--
121
5.15
9.29
K2-V2 (low)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
8*
--
--
--
--
Kimi Linear 48B A3B Instruct
1M
KimiKimi
8*
--
--
--
--
Llama 3.1 405B
128k
MetaMeta
8*
--
--
--
--
LFM2.5-8B-A1B
32.8k
Liquid AILiquid AI
8*
--
324
1.92
9.64
Ring-flash-2.0
128k
InclusionAIInclusionAI
8*
--
--
--
--
Olmo 3.1 32B Think
65.5k
Allen Institute for AIAllen Institute for AI
8*
--
--
--
--
Qwen3.5 2B
262k
AlibabaAlibaba
7
--
--
--
--
Command A
256k
CohereCohere
7*
--
61
1.64
9.82
Llama 3.1 Nemotron 70B
128k
NVIDIANVIDIA
7*
--
114
5.35
9.75
NVIDIA Nemotron 3 Nano
1M
NVIDIANVIDIA
7*
--
142
0.51
4.04
NVIDIA Nemotron Nano 9B V2
131k
NVIDIANVIDIA
7*
--
153
1.86
5.12
Ministral 3 3B
256k
MistralMistral
7
$0.13
202
0.64
3.12
Hermes 4 70B
128k
Nous ResearchNous Research
7*
--
92
1.35
6.76
Granite 4.1 8B
131k
IBMIBM
6*
--
86
0.83
6.63
Sarvam 30B (high)
65.5k
SarvamSarvam
6*
--
--
--
--
Phi-4 Mini
128k
MicrosoftMicrosoft
6
$0.00
43
0.84
12.37
Olmo 3.1 32B Instruct
65.5k
Allen Institute for AIAllen Institute for AI
6*
--
--
--
--
Gemma 4 E2B
128k
GoogleGoogle
6*
--
--
--
--
R1 1776
128k
PerplexityPerplexity
6*
--
--
--
--
Llama 3.2 90B (Vision)
128k
MetaMeta
6*
--
--
--
--
EXAONE 4.0 32B
131k
LG AI ResearchLG AI Research
6*
--
--
--
--
Qwen3.5 2B
262k
AlibabaAlibaba
6
--
--
--
--
Qwen3.5 0.8B
262k
AlibabaAlibaba
5
--
--
--
--
DeepHermes 3 - Mistral 24B
32k
Nous ResearchNous Research
5*
--
--
--
--
Jamba 1.7 Large
256k
AI21 LabsAI21 Labs
5*
--
55
1.33
10.47
Granite 4.0 H Small
128k
IBMIBM
5*
--
44
11.57
22.97
Qwen3 Omni 30B A3B
65.5k
AlibabaAlibaba
5*
--
95
1.88
7.17
Granite 4.1 3B
131k
IBMIBM
5
--
--
--
--
LFM2 24B A2B
32.8k
Liquid AILiquid AI
5*
--
--
--
--
Phi-4
16k
MicrosoftMicrosoft
5*
--
42
1.90
13.87
Nova Micro
130k
AmazonAmazon
4*
--
310
0.94
2.55
MiniCPM-V 4.6 1.3B
262k
OpenBMBOpenBMB
4
--
--
--
--
NVIDIA Nemotron Nano 12B v2 VL
128k
NVIDIANVIDIA
4*
--
182
1.86
4.60
Phi-4 Multimodal
128k
MicrosoftMicrosoft
4*
--
18
0.82
29.12
Jamba Reasoning 3B
262k
AI21 LabsAI21 Labs
4*
--
--
--
--
Reka Flash 3
128k
Reka AIReka AI
4*
--
--
--
--
Olmo 3 7B Think
65.5k
Allen Institute for AIAllen Institute for AI
4*
--
--
--
--
Qwen3.5 0.8B
262k
AlibabaAlibaba
3
--
--
--
--
Molmo 7B-D
4.1k
Allen Institute for AIAllen Institute for AI
3*
--
--
--
--
Llama 3.2 11B (Vision)
128k
MetaMeta
3*
--
31
1.22
17.46
Exaone 4.0 1.2B
64k
LG AI ResearchLG AI Research
3*
--
--
--
--
Olmo 3 7B
65.5k
Allen Institute for AIAllen Institute for AI
2*
--
--
--
--
Exaone 4.0 1.2B
64k
LG AI ResearchLG AI Research
2*
--
--
--
--
LFM2.5-1.2B-Thinking
32k
Liquid AILiquid AI
2*
--
--
--
--
Jamba 1.7 Mini
258k
AI21 LabsAI21 Labs
2*
--
--
--
--
LFM2 2.6B
32.8k
Liquid AILiquid AI
2*
--
--
--
--
LFM2.5-1.2B-Instruct
32k
Liquid AILiquid AI
2*
--
--
--
--
Granite 4.0 H 1B
128k
IBMIBM
2*
--
--
--
--
Gemma 3 270M
32k
GoogleGoogle
2*
--
--
--
--
Apertus 70B Instruct
65.5k
Swiss AI InitiativeSwiss AI Initiative
2*
--
--
--
--
Granite 4.0 Micro
128k
IBMIBM
2*
--
--
--
--
DeepHermes 3 - Llama-3.1 8B
128k
Nous ResearchNous Research
2*
--
--
--
--
Granite 4.0 1B
128k
IBMIBM
2*
--
--
--
--
Molmo2-8B
36.9k
Allen Institute for AIAllen Institute for AI
2*
--
--
--
--
LFM2 8B A1B
32.8k
Liquid AILiquid AI
1*
--
--
--
--
LFM2.5-VL-1.6B
32k
Liquid AILiquid AI
1*
--
351
1.08
2.51
Granite 4.0 350M
32.8k
IBMIBM
1*
--
--
--
--
Tiny Aya Global
8.19k
CohereCohere
1*
--
--
--
--
Apertus 8B Instruct
65.5k
Swiss AI InitiativeSwiss AI Initiative
1*
--
--
--
--
Granite 4.0 H 350M
32.8k
IBMIBM
1*
--
--
--
--
Claude Sonnet 5 (low)
1M
AnthropicAnthropic
--
--
69
2.64
9.92
EXAONE 4.5 33B
262k
LG AI ResearchLG AI Research
--
--
--
--
--
Claude Sonnet 5 (xhigh)
1M
AnthropicAnthropic
--
--
82
42.83
48.94
Gemini 3 Deep Think
128k
GoogleGoogle
--
--
--
--
--
Claude Sonnet 5 (high)
1M
AnthropicAnthropic
--
--
72
8.97
15.87
GPT-5.5 Pro (xhigh)
922k
OpenAIOpenAI
--
--
--
--
--
Cogito v2.1
128k
Deep CogitoDeep Cogito
--
--
--
--
--
Claude Sonnet 5 (medium)
1M
AnthropicAnthropic
--
--
70
3.82
10.97

Key definitions

Maximum number of combined input & output tokens. Output tokens commonly have a significantly lower limit (varied by model).

Frequently Asked Questions

Claude Opus 5 (Adaptive Reasoning, Max Effort) currently ranks #1 on the Artificial Analysis LLM Leaderboard with an Intelligence Index score of 63, out of 179 models ranked.

The top models by Intelligence Index are: 1. Claude Opus 5 (Adaptive Reasoning, Max Effort) (63), 2. Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) (63), 3. Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) (62), 4. Claude Opus 5 (Adaptive Reasoning, High Effort) (61), 5. GPT-5.6 Sol (max) (61).

Celeris-1 is the fastest at 1,548.6 tokens per second, followed by Mercury 2 (1,061.0 t/s) and Ling 3.0 Flash (414.7 t/s).

GPT-5.6 Luna (low) has the lowest cost per Intelligence Index task at $0.01, followed by MiMo-V2.5 ($0.01) and Llama 4 Scout ($0.01).

Kimi K3 (max) is the highest-ranked open weights model with an Intelligence Index score of 60. There are 100 open weights models out of 179 total on the leaderboard.

The top open weights models by Intelligence Index are: 1. Kimi K3 (max) (60), 2. GLM-5.2 (max) (53), 3. DeepSeek V4 Flash 0731 (Reasoning, Max Effort) (52).

Claude Opus 5 (Adaptive Reasoning, Max Effort) leads among 134 reasoning models with an Intelligence Index score of 63. Reasoning models use extended thinking to solve complex problems before responding.

The leaderboard includes filters to narrow results by model type (reasoning vs non-reasoning), openness (open weights vs proprietary), and other criteria. You can also adjust prompt options to see how performance varies with different input lengths.

Click on any model name in the leaderboard to visit its dedicated comparison page with detailed charts covering intelligence, pricing, speed, latency, and more. You can also compare API providers for each model. View all models