LLM Leaderboard - Comparison of AI models from OpenAI, Anthropic, Google, SpaceXAI & others

Comparison and ranking the performance of over 250 AI models (LLMs) across key metrics including intelligence, price, performance and speed (output speed - tokens per second & latency - TTFT), context window & others.

For more details including relating to our methodology, see our FAQs.

Intelligence

Updated

Claude Fable 5 (with fallback) and GPT-5.6 Sol (max) are the highest intelligence models, followed by GPT-5.6 Sol (xhigh) and Kimi K3.

Output Speed

Mercury 2 and LFM2.5-VL-1.6B are the fastest models, followed by Granite 4.0 H Small and Step 3.7 Flash.

Latency

North Mini Code and Gemini 2.5 Flash-Lite are the lowest latency models, followed by Command A+ and Gemini 2.5 Flash.

Cost per Task

Llama 4 Scout and MiMo-V2.5 have the lowest cost per task, followed by gpt-oss-120b (low) and gpt-oss-20b (high).

Context Window

Llama 4 Scout and Grok 4.20 0309 support the largest context windows, followed by Gemini 1.5 Pro (May) and Grok 4.1 Fast.

Further Analysis
Claude Fable 5 (with fallback)
1M
AnthropicAnthropic
60
$2.75
68
111.29
118.61
GPT-5.6 Sol (max)
1M
OpenAIOpenAI
59
$1.04
63
145.73
153.63
GPT-5.6 Sol (xhigh)
1M
OpenAIOpenAI
58
$0.68
61
57.83
66.06
Kimi K3
1.05M
KimiKimi
57
$0.95
34
5.50
78.55
GPT-5.6 Sol (high)
1M
OpenAIOpenAI
56
$0.45
59
11.42
19.94
Claude Opus 4.8 (max)
1M
AnthropicAnthropic
56
$1.80
59
27.04
35.50
GPT-5.6 Terra (max)
1M
OpenAIOpenAI
55
$0.82
135
144.90
148.59
Grok 4.5 (high)
500k
SpaceXAISpaceXAI
54
$0.31
69
10.00
17.27
GPT-5.6 Sol (medium)
1M
OpenAIOpenAI
54
$0.31
57
6.58
15.33
Claude Sonnet 5 (max)
1M
AnthropicAnthropic
53
$1.53
84
179.03
184.99
GPT-5.6 Terra (xhigh)
1M
OpenAIOpenAI
52
$0.48
130
10.63
14.49
GPT-5.6 Luna (max)
1M
OpenAIOpenAI
51
$0.21
190
103.19
105.82
GLM-5.2 (max)
1M
Z AIZ AI
51
$0.47
201
1.36
13.78
Muse Spark 1.1 (xhigh)
1.05M
MetaMeta
51
$0.26
124
1.25
21.44
Gemini 3.5 Flash
1M
GoogleGoogle
50
$0.59
165
20.22
23.24
Gemini 3.6 Flash
1M
GoogleGoogle
50
$0.50
304
11.54
13.19
GPT-5.6 Sol (low)
1M
OpenAIOpenAI
49
$0.20
54
3.21
12.51
GPT-5.6 Luna (xhigh)
1M
OpenAIOpenAI
49
$0.14
190
38.69
41.32
GPT-5.6 Terra (high)
1M
OpenAIOpenAI
49
$0.34
122
2.95
7.05
Gemini 3.1 Pro Preview
1M
GoogleGoogle
46
$0.29
121
33.44
37.56
GPT-5.6 Luna (high)
1M
OpenAIOpenAI
46
$0.09
184
7.43
10.15
Qwen3.7 Max
1M
AlibabaAlibaba
46
$1.03
205
2.43
16.60
GPT-5.6 Terra (medium)
1M
OpenAIOpenAI
46
$0.18
121
1.35
5.48
Gemini 3.5 Flash (medium)
1M
GoogleGoogle
45*
--
188
19.59
22.25
MiniMax-M3
1M
MiniMaxMiniMax
44
$0.12
105
1.28
25.16
DeepSeek V4 Pro (max)
1M
DeepSeekDeepSeek
44
$0.04
63
1.73
79.24
GPT-5.3 Codex (xhigh)
400k
OpenAIOpenAI
44*
--
141
68.72
72.27
Motif 3 (Beta)
262k
Motif TechnologiesMotif Technologies
44
--
--
--
--
DeepSeek V4 Pro (high)
1M
DeepSeekDeepSeek
43
$0.04
67
1.81
38.95
Muse Spark
262k
MetaMeta
43
--
--
--
--
Claude Opus 4.7 (Non-reasoning, high)
1M
AnthropicAnthropic
43*
--
47
1.66
12.36
MiMo-V2.5-Pro
1M
XiaomiXiaomi
42
$0.03
60
3.00
44.79
Kimi K2.7 Code
256k
KimiKimi
42
--
48
2.94
59.92
Claude Sonnet 5 (Non-reasoning)
1M
AnthropicAnthropic
42
$0.37
67
1.23
8.69
Hy3
256k
TencentTencent
41
$0.00
61
2.44
43.74
GPT-5.6 Sol (Non-reasoning)
1M
OpenAIOpenAI
41
$0.20
53
0.95
10.31
Nex-N2-Pro
262k
Nex AGINex AGI
41
--
137
1.69
19.95
Inkling
1M
Thinking MachinesThinking Machines
41
--
--
--
--
GPT-5.6 Terra (low)
1M
OpenAIOpenAI
40
$0.15
124
1.29
5.33
DeepSeek V4 Flash (max)
1M
DeepSeekDeepSeek
40
$0.02
121
1.27
51.78
Qwen3.6 Plus
1M
AlibabaAlibaba
40
$0.31
53
2.63
116.66
Qwen3.7 Plus
1M
AlibabaAlibaba
39
$0.21
53
2.98
49.87
JT-4.1 Flash 236B A21B
256k
China MobileChina Mobile
39
--
--
--
--
GPT-5.6 Luna (medium)
1M
OpenAIOpenAI
38
$0.05
192
1.75
4.35
Nemotron 3 Ultra
262k
NVIDIANVIDIA
38
$0.24
201
1.19
15.01
DeepSeek V4 Flash (high)
1M
DeepSeekDeepSeek
37
$0.04
--
--
--
MiMo-V2.5
1M
XiaomiXiaomi
37
$0.01
62
3.81
43.91
Qwen3.6 27B
262k
AlibabaAlibaba
37
$0.27
57
3.78
112.51
Gemini 3.5 Flash-Lite
1M
GoogleGoogle
36
$0.09
350
7.24
8.67
MiMo-V2-Omni-0327
256k
XiaomiXiaomi
36*
--
--
--
--
Grok 4.3 (medium)
1M
SpaceXAISpaceXAI
36*
--
101
10.37
15.32
Grok 4.3 (low)
1M
SpaceXAISpaceXAI
35*
--
109
5.74
10.32
MiMo-V2-Omni
256k
XiaomiXiaomi
35*
--
--
--
--
Gemini 3.5 Flash (minimal)
1M
GoogleGoogle
35*
--
169
0.95
3.91
Kimi K2.6
256k
KimiKimi
35*
--
47
2.93
13.64
Claude Sonnet 4.6 (Non-reasoning, Low Effort)
1M
AnthropicAnthropic
34*
--
44
1.27
12.71
GLM-5.2
1M
Z AIZ AI
34
--
130
1.62
5.46
GPT-5.6 Terra (Non-reasoning)
1M
OpenAIOpenAI
34
$0.18
122
0.67
4.76
KAT-Coder-Pro V2
256k
KwaiKATKwaiKAT
34
--
--
--
--
Qwen3.5 397B A17B
262k
AlibabaAlibaba
34
$0.33
63
2.59
61.24
Hy3-preview
256k
TencentTencent
34*
--
151
3.00
19.51
LongCat 2.0
1M
LongCatLongCat
33
--
--
--
--
GPT-5.6 Luna (low)
1M
OpenAIOpenAI
33
$0.04
179
1.23
4.02
MiMo-V2-Flash (Feb 2026)
256k
XiaomiXiaomi
33*
--
--
--
--
Qwen3.5 122B A10B
262k
AlibabaAlibaba
32
$0.24
133
2.38
21.12
Qwen3.5 397B A17B
262k
AlibabaAlibaba
32*
--
63
2.47
10.40
Qwen3.6 35B A3B
262k
AlibabaAlibaba
32
$0.18
149
2.27
41.87
DeepSeek V4 Pro
1M
DeepSeekDeepSeek
31*
--
68
1.79
9.12
Qwen3.5 Omni Plus
256k
AlibabaAlibaba
31*
--
54
2.34
11.64
Ring-2.6-1T
262k
InclusionAIInclusionAI
31
$0.35
129
3.23
22.58
Qwen3.6 27B
262k
AlibabaAlibaba
30
$0.36
57
3.59
12.37
o3
200k
OpenAIOpenAI
30*
--
159
7.60
10.74
Step 3.7 Flash
262k
StepFunStepFun
30
--
391
0.85
7.25
Mistral Medium 3.5
256k
MistralMistral
30
$0.56
123
1.95
22.20
Claude 4.5 Haiku
200k
AnthropicAnthropic
30
$0.24
97
16.21
21.37
Gemma 4 31B
256k
GoogleGoogle
29
--
35
1.12
64.41
GPT-5.5 Instant (June 2026)
400k
OpenAIOpenAI
29
$0.54
--
--
--
DeepSeek V4 Flash
1M
DeepSeekDeepSeek
29*
--
118
1.34
5.57
JT-35B-Flash
256k
China MobileChina Mobile
28*
--
--
--
--
KAT-Coder-Pro V1
256k
KwaiKATKwaiKAT
28*
--
--
--
--
MiMo-V2.5-Pro
1M
XiaomiXiaomi
28*
--
58
2.25
10.87
Qwen3.5 122B A10B
262k
AlibabaAlibaba
28
$0.18
149
2.42
5.77
GPT-5.6 Luna (Non-reasoning)
1M
OpenAIOpenAI
27
$0.05
199
0.63
3.15
Hy3-preview
256k
TencentTencent
26*
--
128
3.04
6.93
Ling-2.6-1T
262k
InclusionAIInclusionAI
26*
--
--
--
--
Step 3.5 Flash 2603
256k
StepFunStepFun
26*
--
291
1.00
9.60
Doubao Seed Code
256k
ByteDance SeedByteDance Seed
26*
--
--
--
--
Gemini 2.5 Pro
1M
GoogleGoogle
26
$0.20
142
21.64
25.17
Gemma 4 26B A4B
256k
GoogleGoogle
26
$0.03
--
--
--
NVIDIA Nemotron 3 Super
1M
NVIDIANVIDIA
25
$0.21
209
1.10
13.05
Gemini 3.1 Flash-Lite
1M
GoogleGoogle
25
$0.04
333
5.85
7.35
Grok 4.3 (Non-reasoning)
1M
SpaceXAISpaceXAI
25
$0.29
98
0.98
6.09
K-EXAONE
256k
LG AI ResearchLG AI Research
25*
--
--
--
--
MiMo-V2-Flash
256k
XiaomiXiaomi
25
--
--
--
--
Trinity Large Thinking
512k
Arcee AIArcee AI
24*
--
199
1.00
13.59
Qwen3.6 35B A3B
262k
AlibabaAlibaba
24
$0.60
191
2.27
4.88
Qwen3.5 35B A3B
262k
AlibabaAlibaba
24
$0.23
145
2.13
5.59
gpt-oss-120b (high)
131k
OpenAIOpenAI
24
$0.06
303
0.85
9.09
Claude 4.5 Haiku
200k
AnthropicAnthropic
24*
--
94
0.92
6.26
EXAONE 4.5 33B
262k
LG AI ResearchLG AI Research
23*
--
--
--
--
Command A+
192k
CohereCohere
23
$0.00
177
0.42
14.58
Gemma 4 12B
256k
GoogleGoogle
22*
--
103
2.55
26.77
ERNIE 5.0 Thinking Preview
128k
BaiduBaidu
22*
--
--
--
--
Gemma 4 31B
256k
GoogleGoogle
22
$0.03
66
1.89
9.46
Nova 2.0 Pro Preview (medium)
256k
AmazonAmazon
22
$0.17
129
11.67
31.03
Qwen3.5 9B
262k
AlibabaAlibaba
21
$0.16
58
1.82
45.16
Mercury 2
128k
InceptionInception
21
$0.07
982
4.37
4.88
Nemotron Cascade 2 30B A3B
1M
NVIDIANVIDIA
21*
--
--
--
--
Qwen3 Coder Next
256k
AlibabaAlibaba
21
$0.33
107
1.18
5.87
Nova 2.0 Omni (medium)
1M
AmazonAmazon
21*
--
--
--
--
Apriel-v1.6-15B-Thinker
128k
ServiceNowServiceNow
21*
--
--
--
--
Qwen3.5 9B
262k
AlibabaAlibaba
20*
--
--
--
--
Gemma 4 26B A4B
256k
GoogleGoogle
20*
--
50
1.12
11.21
Qwen3.5 4B
262k
AlibabaAlibaba
20*
--
23
0.79
110.92
North Mini Code
256k
CohereCohere
20
$0.00
100
0.34
25.24
Nova 2.0 Pro Preview (low)
256k
AmazonAmazon
20
$0.21
126
11.33
31.19
Mistral Small 4
256k
MistralMistral
20
$0.10
159
0.74
16.43
Devstral 2
256k
MistralMistral
19
$0.00
68
1.37
8.75
Nova 2.0 Lite (medium)
1M
AmazonAmazon
19*
--
147
13.52
30.57
Qwen3.5 Omni Flash
256k
AlibabaAlibaba
19*
--
247
1.82
3.84
JT-MINI
128k
China MobileChina Mobile
19*
--
--
--
--
Nova 2.0 Lite (high)
1M
AmazonAmazon
18
$0.25
143
19.12
36.60
Magistral Medium 1.2
128k
MistralMistral
18
$0.75
42
1.75
61.70
Nova 2.0 Lite (low)
1M
AmazonAmazon
18*
--
150
9.58
26.28
HyperNova 60B 2605
131k
Multiverse ComputingMultiverse Computing
18
$0.02
360
0.80
7.74
Devstral Small 2
256k
MistralMistral
17
$0.00
70
1.40
8.57
K2 Think V2
262k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
17
--
--
--
--
LongCat Flash Lite
256k
LongCatLongCat
17*
--
--
--
--
HyperCLOVA X SEED Think (32B)
128k
NaverNaver
17*
--
--
--
--
K-EXAONE
256k
LG AI ResearchLG AI Research
17*
--
--
--
--
Qwen3 Next 80B A3B
262k
AlibabaAlibaba
17
$0.17
179
2.10
16.10
Nova 2.0 Omni (low)
1M
AmazonAmazon
17*
--
--
--
--
Mi:dm K 2.5 Pro
128k
Korea TelecomKorea Telecom
16*
--
--
--
--
Qwen3.5 4B
262k
AlibabaAlibaba
16*
--
15
0.97
33.41
Mistral Large 3
256k
MistralMistral
16
$0.06
51
1.12
10.95
INTELLECT-3
131k
Prime IntellectPrime Intellect
16*
--
--
--
--
Solar Open 100B
128k
UpstageUpstage
15*
--
--
--
--
Nemotron 3 Nano Omni 30B A3B Reasoning
256k
NVIDIANVIDIA
15*
--
315
0.98
8.91
gpt-oss-120b (low)
131k
OpenAIOpenAI
15
$0.02
326
0.87
8.55
gpt-oss-20b (high)
131k
OpenAIOpenAI
15
$0.02
209
0.79
12.75
Nova 2.0 Pro Preview
256k
AmazonAmazon
14
$0.25
109
1.07
5.67
gpt-oss-20b (low)
131k
OpenAIOpenAI
14*
--
240
0.85
11.28
Llama 4 Maverick
1M
MetaMeta
14
$0.03
110
0.95
5.52
K2-V2 (high)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
14*
--
--
--
--
NVIDIA Nemotron 3 Nano
1M
NVIDIANVIDIA
14
$0.02
156
1.19
17.27
Solar Pro 3
128k
UpstageUpstage
14
--
--
--
--
Ling 2.6 Flash
262k
InclusionAIInclusionAI
14
--
172
1.13
4.04
Qwen3 Next 80B A3B
262k
AlibabaAlibaba
14*
--
183
2.13
4.86
Tri-21B-think Preview
32k
Trillion LabsTrillion Labs
14*
--
--
--
--
DiffusionGemma 26B A4B
256k
GoogleGoogle
13
--
--
--
--
Gemma 4 12B (Non-reasoning)
262k
GoogleGoogle
13*
--
98
2.44
7.53
Motif-2-12.7B
128k
Motif TechnologiesMotif Technologies
13*
--
--
--
--
Nova Premier
1M
AmazonAmazon
13*
--
29
2.97
20.18
Gemma 4 E4B
128k
GoogleGoogle
12*
--
95
1.12
27.51
K2-V2 (medium)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
12*
--
--
--
--
Llama Nemotron Super 49B v1.5
128k
NVIDIANVIDIA
12*
--
106
3.79
27.45
Mistral Small 4
256k
MistralMistral
12*
--
150
0.78
4.12
Tri-21B-Think
32k
Trillion LabsTrillion Labs
12*
--
--
--
--
MiniCPM5-1B
128k
OpenBMBOpenBMB
12*
--
--
--
--
Sarvam 105B (high)
128k
SarvamSarvam
12*
--
--
--
--
Nova 2.0 Lite
1M
AmazonAmazon
12*
--
143
1.05
4.54
MiniCPM5-1B
128k
OpenBMBOpenBMB
12*
--
--
--
--
Magistral Small 1.2
128k
MistralMistral
11
$0.25
81
0.93
31.63
Nanbeige4.1-3B
256k
NanbeigeNanbeige
11
--
--
--
--
Ministral 3 14B
256k
MistralMistral
11
$0.15
64
0.86
8.68
EXAONE 4.0 32B
131k
LG AI ResearchLG AI Research
11*
--
--
--
--
Nova 2.0 Omni
1M
AmazonAmazon
11*
--
--
--
--
Llama 4 Scout
10M
MetaMeta
10
$0.01
79
0.81
7.14
Hermes 4 70B
128k
Nous ResearchNous Research
10*
--
91
1.38
28.86
Falcon-H1R-7B
256k
TII UAETII UAE
10*
--
--
--
--
Qwen3 Omni 30B A3B
65.5k
AlibabaAlibaba
10*
--
100
1.91
26.89
Step3 VL 10B
65.5k
StepFunStepFun
9*
--
--
--
--
Llama 3.3 70B
128k
MetaMeta
9
$0.08
81
1.71
7.85
Gemma 4 E2B
128k
GoogleGoogle
9*
--
--
--
--
Llama Nemotron Ultra
128k
NVIDIANVIDIA
9*
--
53
2.33
49.70
ERNIE 4.5 300B A47B
131k
BaiduBaidu
9*
--
--
--
--
Hermes 4 405B
128k
Nous ResearchNous Research
9*
--
37
2.38
69.84
Solar Pro 2
65.5k
UpstageUpstage
9*
--
--
--
--
NVIDIA Nemotron Nano 12B v2 VL
128k
NVIDIANVIDIA
9*
--
103
5.42
29.69
Ministral 3 8B
256k
MistralMistral
9
$0.18
113
0.75
5.17
Gemma 4 E4B
128k
GoogleGoogle
9*
--
95
1.10
6.36
Granite 4.1 30B
131k
IBMIBM
9
--
--
--
--
NVIDIA Nemotron Nano 9B V2
131k
NVIDIANVIDIA
9*
--
73
8.91
43.24
Hermes 4 405B
128k
Nous ResearchNous Research
9*
--
35
2.41
16.65
NVIDIA Nemotron 3 Nano 4B
262k
NVIDIANVIDIA
9
--
--
--
--
Llama Nemotron Super 49B v1.5
128k
NVIDIANVIDIA
9*
--
120
4.68
8.86
K2-V2 (low)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
9*
--
--
--
--
Kimi Linear 48B A3B Instruct
1M
KimiKimi
9*
--
--
--
--
Llama 3.1 405B
128k
MetaMeta
9*
--
--
--
--
LFM2.5-8B-A1B
32.8k
Liquid AILiquid AI
8*
--
336
2.92
10.36
Ring-flash-2.0
128k
InclusionAIInclusionAI
8*
--
--
--
--
Olmo 3.1 32B Think
65.5k
Allen Institute for AIAllen Institute for AI
8*
--
--
--
--
Solar Pro 2
65.5k
UpstageUpstage
8*
--
--
--
--
Command A
256k
CohereCohere
8*
--
62
1.68
9.74
Qwen3.5 2B
262k
AlibabaAlibaba
8
--
--
--
--
Llama 3.1 Nemotron 70B
128k
NVIDIANVIDIA
8*
--
108
4.29
8.91
NVIDIA Nemotron 3 Nano
1M
NVIDIANVIDIA
7*
--
103
0.92
5.79
NVIDIA Nemotron Nano 9B V2
131k
NVIDIANVIDIA
7*
--
102
2.72
7.64
Hermes 4 70B
128k
Nous ResearchNous Research
7*
--
92
1.38
6.81
Ministral 3 3B
256k
MistralMistral
7
$0.13
164
0.61
3.66
Granite 4.1 8B
131k
IBMIBM
7*
--
118
0.79
5.02
Sarvam 30B (high)
65.5k
SarvamSarvam
7*
--
--
--
--
Olmo 3.1 32B Instruct
65.5k
Allen Institute for AIAllen Institute for AI
6*
--
--
--
--
Gemma 4 E2B
128k
GoogleGoogle
6*
--
--
--
--
R1 1776
128k
PerplexityPerplexity
6*
--
--
--
--
Llama 3.2 90B (Vision)
128k
MetaMeta
6*
--
--
--
--
Phi-4 Mini
128k
MicrosoftMicrosoft
6
$0.00
44
0.82
12.07
EXAONE 4.0 32B
131k
LG AI ResearchLG AI Research
6*
--
--
--
--
Qwen3.5 2B
262k
AlibabaAlibaba
6
--
--
--
--
Qwen3.5 0.8B
262k
AlibabaAlibaba
5
--
--
--
--
DeepHermes 3 - Mistral 24B
32k
Nous ResearchNous Research
5*
--
--
--
--
Jamba 1.7 Large
256k
AI21 LabsAI21 Labs
5*
--
57
1.63
10.44
Granite 4.0 H Small
128k
IBMIBM
5*
--
401
10.26
11.50
Qwen3 Omni 30B A3B
65.5k
AlibabaAlibaba
5*
--
93
1.85
7.25
LFM2 24B A2B
32.8k
Liquid AILiquid AI
5*
--
--
--
--
Phi-4
16k
MicrosoftMicrosoft
5*
--
--
--
--
Nova Micro
130k
AmazonAmazon
5*
--
283
0.88
2.64
Granite 4.1 3B
131k
IBMIBM
5
--
--
--
--
NVIDIA Nemotron Nano 12B v2 VL
128k
NVIDIANVIDIA
5*
--
200
1.25
3.75
Phi-4 Multimodal
128k
MicrosoftMicrosoft
5*
--
18
0.79
28.36
MiniCPM-V 4.6 1.3B
262k
OpenBMBOpenBMB
4
--
--
--
--
Jamba Reasoning 3B
262k
AI21 LabsAI21 Labs
4*
--
--
--
--
Reka Flash 3
128k
Reka AIReka AI
4*
--
--
--
--
Olmo 3 7B Think
65.5k
Allen Institute for AIAllen Institute for AI
4*
--
--
--
--
Molmo 7B-D
4.1k
Allen Institute for AIAllen Institute for AI
4*
--
--
--
--
Ling-mini-2.0
131k
InclusionAIInclusionAI
4*
--
--
--
--
Llama 3.2 11B (Vision)
128k
MetaMeta
3*
--
17
1.49
30.58
Qwen3.5 0.8B
262k
AlibabaAlibaba
3
--
--
--
--
Exaone 4.0 1.2B
64k
LG AI ResearchLG AI Research
3*
--
--
--
--
Olmo 3 7B
65.5k
Allen Institute for AIAllen Institute for AI
3*
--
--
--
--
Exaone 4.0 1.2B
64k
LG AI ResearchLG AI Research
3*
--
--
--
--
LFM2.5-1.2B-Thinking
32k
Liquid AILiquid AI
3*
--
--
--
--
Jamba 1.7 Mini
258k
AI21 LabsAI21 Labs
3*
--
--
--
--
LFM2 2.6B
32.8k
Liquid AILiquid AI
3*
--
--
--
--
LFM2.5-1.2B-Instruct
32k
Liquid AILiquid AI
3*
--
--
--
--
Granite 4.0 H 1B
128k
IBMIBM
3*
--
--
--
--
Gemma 3 270M
32k
GoogleGoogle
2*
--
--
--
--
Apertus 70B Instruct
65.5k
Swiss AI InitiativeSwiss AI Initiative
2*
--
--
--
--
Granite 4.0 Micro
128k
IBMIBM
2*
--
--
--
--
DeepHermes 3 - Llama-3.1 8B
128k
Nous ResearchNous Research
2*
--
--
--
--
Granite 4.0 1B
128k
IBMIBM
2*
--
--
--
--
Molmo2-8B
36.9k
Allen Institute for AIAllen Institute for AI
2*
--
--
--
--
LFM2 8B A1B
32.8k
Liquid AILiquid AI
2*
--
--
--
--
LFM2.5-VL-1.6B
32k
Liquid AILiquid AI
1*
--
414
2.36
3.56
Granite 4.0 350M
32.8k
IBMIBM
1*
--
--
--
--
Tiny Aya Global
8.19k
CohereCohere
1*
--
--
--
--
Apertus 8B Instruct
65.5k
Swiss AI InitiativeSwiss AI Initiative
1*
--
--
--
--
Granite 4.0 H 350M
32.8k
IBMIBM
1*
--
--
--
--
Claude Sonnet 5 (low)
1M
AnthropicAnthropic
--
--
62
1.57
9.59
EXAONE 4.5 33B
262k
LG AI ResearchLG AI Research
--
--
--
--
--
Claude Sonnet 5 (xhigh)
1M
AnthropicAnthropic
--
--
74
21.86
28.64
Gemini 3 Deep Think
128k
GoogleGoogle
--
--
--
--
--
Claude Sonnet 5 (high)
1M
AnthropicAnthropic
--
--
70
5.93
13.07
Mi:dm K 2.5 Pro Preview
128k
Korea TelecomKorea Telecom
--
--
--
--
--
GPT-5.5 Pro (xhigh)
922k
OpenAIOpenAI
--
--
--
--
--
Cogito v2.1
128k
Deep CogitoDeep Cogito
--
--
--
--
--
Claude Sonnet 5 (medium)
1M
AnthropicAnthropic
--
--
63
1.85
9.84

Key definitions

Maximum number of combined input & output tokens. Output tokens commonly have a significantly lower limit (varied by model).

Tokens per second received while the model is generating tokens (ie. after first chunk has been received from the API for models which support streaming).

Time to first token received, in seconds, after API request sent. For reasoning models which share reasoning tokens, this will be the first reasoning token. For models which do not support streaming, this represents time to receive the completion.

Average cost per task in the index. Costs are split by input, cache hit, cache write, reasoning, and answer token pricing where canonical token counts are available.

Price per token included in the request/message sent to the API, represented as USD per million Tokens.

Price per token for cached prompts (previously processed), typically offering a significant discount compared to regular input price, represented as USD per million tokens. The values shown here are the cache hit price; cache write and cache storage are billed separately and vary by provider — see "Cache pricing by provider" for detail.

Price per token to write prompt tokens into the cache so that later requests can hit them, represented as USD per million tokens. Some providers charge a premium over the standard input price to create a cache entry (e.g. Anthropic), while others cache automatically with no separate write fee.

Price per token generated by the model (received from the API), represented as USD per million Tokens.

Metrics are 'live' and are based on the past 72 hours of measurements, measurements are taken 8 times a day for single requests and 2 times per day for parallel requests.

Frequently Asked Questions

Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) currently ranks #1 on the Artificial Analysis LLM Leaderboard with an Intelligence Index score of 60, out of 156 models ranked.

The top models by Intelligence Index are: 1. Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) (60), 2. GPT-5.6 Sol (max) (59), 3. GPT-5.6 Sol (xhigh) (58), 4. Kimi K3 (57), 5. GPT-5.6 Sol (high) (56).

Mercury 2 is the fastest at 981.5 tokens per second, followed by LFM2.5-VL-1.6B (414.3 t/s) and Granite 4.0 H Small (400.7 t/s).

Llama 4 Scout has the lowest cost per Intelligence Index task at $0.01, followed by MiMo-V2.5 ($0.01) and gpt-oss-120b (low) ($0.02).

GLM-5.2 (max) is the highest-ranked open weights model with an Intelligence Index score of 51. There are 86 open weights models out of 156 total on the leaderboard.

The top open weights models by Intelligence Index are: 1. GLM-5.2 (max) (51), 2. MiniMax-M3 (44), 3. DeepSeek V4 Pro (Reasoning, Max Effort) (44).

Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) leads among 112 reasoning models with an Intelligence Index score of 60. Reasoning models use extended thinking to solve complex problems before responding.

The leaderboard includes filters to narrow results by model type (reasoning vs non-reasoning), openness (open weights vs proprietary), and other criteria. You can also adjust prompt options to see how performance varies with different input lengths.

Click on any model name in the leaderboard to visit its dedicated comparison page with detailed charts covering intelligence, pricing, speed, latency, and more. You can also compare API providers for each model. View all models