Ranking de modelos de lenguaje (LLM) - Comparación de modelos de IA de OpenAI, Anthropic, Google, SpaceXAI y otros

Comparación y ranking del rendimiento de más de 250 modelos de IA (LLM) en métricas clave como inteligencia, precio, rendimiento y velocidad (velocidad de salida - tokens por segundo y latencia - TTFT), ventana de contexto y otras.

Para más detalles, incluso sobre nuestra metodología, consulta nuestras preguntas frecuentes.

Inteligencia

Claude Opus 5 (max) y Claude Opus 5 (xhigh) son los modelos con mayor inteligencia, seguidos por Claude Fable 5 (with fallback) y GPT-5.6 Sol (max).

Velocidad de salida

Mercury 2 y Gemini 3.5 Flash-Lite son los modelos más rápidos, seguidos por LFM2.5-VL-1.6B y Step 3.7 Flash.

Latencia

Gemini 2.5 Flash-Lite y Command A+ son los modelos con menor latencia, seguidos por North Mini Code y Gemini 2.5 Flash.

Costo por tarea

Llama 4 Scout y MiMo-V2.5 tienen el menor costo por tarea, seguidos por gpt-oss-120b (low) y NVIDIA Nemotron 3 Nano.

Ventana de contexto

Llama 4 Scout y Grok 4.20 0309 admiten las ventanas de contexto más grandes, seguidos por Gemini 1.5 Pro (May) y Grok 4.1 Fast.

Análisis adicional
Claude Opus 5 (max)
1M
AnthropicAnthropic
61
$2.03
55
69.71
78.81
Claude Opus 5 (xhigh)
1M
AnthropicAnthropic
60
$1.56
53
40.62
49.99
Claude Fable 5 (with fallback)
1M
AnthropicAnthropic
60
$2.75
72
184.00
190.93
GPT-5.6 Sol (max)
1M
OpenAIOpenAI
59
$1.54
81
130.03
136.23
Claude Opus 5 (high)
1M
AnthropicAnthropic
59
$1.06
57
22.29
31.10
GPT-5.6 Sol (xhigh)
1M
OpenAIOpenAI
58
$0.95
82
44.90
51.02
Kimi K3
1.05M
KimiKimi
57
$0.72
33
148.54
223.66
Claude Opus 5 (medium)
1M
AnthropicAnthropic
56
$0.62
52
5.98
15.60
GPT-5.6 Sol (high)
1M
OpenAIOpenAI
56
$0.62
78
12.48
18.92
Claude Opus 4.8 (max)
1M
AnthropicAnthropic
56
$1.80
56
30.95
39.87
GPT-5.6 Terra (max)
1M
OpenAIOpenAI
55
$0.82
137
172.72
176.36
Grok 4.5 (high)
500k
SpaceXAISpaceXAI
54
$0.35
58
8.79
17.37
GPT-5.6 Sol (medium)
1M
OpenAIOpenAI
54
$0.41
74
4.92
11.67
Claude Sonnet 5 (max)
1M
AnthropicAnthropic
53
$1.53
74
174.30
181.03
GPT-5.6 Terra (xhigh)
1M
OpenAIOpenAI
52
$0.48
134
20.06
23.79
GPT-5.6 Luna (max)
1M
OpenAIOpenAI
51
$0.28
195
127.85
130.42
GLM-5.2 (max)
1M
Z AIZ AI
51
$0.32
225
1.31
12.41
Muse Spark 1.1 (xhigh)
1.05M
MetaMeta
51
$0.26
129
1.46
20.88
Claude Opus 5 (low)
1M
AnthropicAnthropic
51
$0.36
52
3.66
13.23
Gemini 3.5 Flash
1M
GoogleGoogle
50
$0.59
188
28.40
31.07
Gemini 3.6 Flash
1M
GoogleGoogle
50
$0.50
240
18.83
20.92
GPT-5.6 Sol (low)
1M
OpenAIOpenAI
49
$0.24
72
3.09
10.06
GPT-5.6 Luna (xhigh)
1M
OpenAIOpenAI
49
$0.18
178
31.90
34.71
GPT-5.6 Terra (high)
1M
OpenAIOpenAI
49
$0.34
125
2.16
6.17
Gemini 3.1 Pro Preview
1M
GoogleGoogle
46
$0.29
127
42.32
46.25
GPT-5.6 Luna (high)
1M
OpenAIOpenAI
46
$0.12
185
9.07
11.77
Qwen3.7 Max
1M
AlibabaAlibaba
46
$1.03
201
2.54
16.98
GPT-5.6 Terra (medium)
1M
OpenAIOpenAI
46
$0.18
118
1.62
5.85
Gemini 3.5 Flash (medium)
1M
GoogleGoogle
45*
--
195
23.05
25.61
MiniMax-M3
1M
MiniMaxMiniMax
44
$0.12
83
1.61
31.72
DeepSeek V4 Pro (max)
1M
DeepSeekDeepSeek
44
$0.04
71
1.54
69.85
GPT-5.3 Codex (xhigh)
400k
OpenAIOpenAI
44*
--
138
65.95
69.58
Motif 3 (Beta)
262k
Motif TechnologiesMotif Technologies
44
--
--
--
--
DeepSeek V4 Pro (high)
1M
DeepSeekDeepSeek
43
$0.04
71
1.57
36.45
Muse Spark
262k
MetaMeta
43
--
--
--
--
Claude Opus 4.7 (Non-reasoning, high)
1M
AnthropicAnthropic
43*
--
44
1.90
13.14
MiMo-V2.5-Pro
1M
XiaomiXiaomi
42
$0.06
67
2.75
40.19
Kimi K2.7 Code
256k
KimiKimi
42
--
46
2.79
62.57
Claude Sonnet 5 (Non-reasoning)
1M
AnthropicAnthropic
42
$0.37
61
1.16
9.32
Hy3
256k
TencentTencent
41
$0.03
58
2.49
45.65
GPT-5.6 Sol (Non-reasoning)
1M
OpenAIOpenAI
41
$0.26
73
0.99
7.87
Nex-N2-Pro
262k
Nex AGINex AGI
41
--
142
1.77
19.38
Inkling
1M
Thinking MachinesThinking Machines
41
--
87
1.75
30.34
GPT-5.6 Terra (low)
1M
OpenAIOpenAI
40
$0.15
129
1.39
5.26
DeepSeek V4 Flash (max)
1M
DeepSeekDeepSeek
40
$0.02
122
1.25
51.49
Qwen3.6 Plus
1M
AlibabaAlibaba
40
$0.31
52
2.51
118.81
Qwen3.7 Plus
1M
AlibabaAlibaba
39
$0.22
53
2.93
49.84
JT-4.1 Flash 236B A21B
256k
China MobileChina Mobile
39
--
--
--
--
Agnes 2.5 Pro Alpha
1M
Sapiens AISapiens AI
39
--
107
1.66
24.98
GPT-5.6 Luna (medium)
1M
OpenAIOpenAI
38
$0.07
192
2.84
5.45
Nemotron 3 Ultra
262k
NVIDIANVIDIA
38
$0.25
204
1.15
14.77
DeepSeek V4 Flash (high)
1M
DeepSeekDeepSeek
37
$0.04
--
--
--
MiMo-V2.5
1M
XiaomiXiaomi
37
$0.01
67
4.39
41.66
Qwen3.6 27B
262k
AlibabaAlibaba
37
$0.27
56
3.67
114.46
Gemini 3.5 Flash-Lite
1M
GoogleGoogle
36
$0.09
433
8.38
9.54
MiMo-V2-Omni-0327
256k
XiaomiXiaomi
36*
--
--
--
--
Grok 4.3 (medium)
1M
SpaceXAISpaceXAI
36*
--
110
12.58
17.12
Grok 4.3 (low)
1M
SpaceXAISpaceXAI
35*
--
105
5.85
10.60
MiMo-V2-Omni
256k
XiaomiXiaomi
35*
--
--
--
--
Gemini 3.5 Flash (minimal)
1M
GoogleGoogle
35*
--
167
0.92
3.91
Kimi K2.6
256k
KimiKimi
35*
--
38
2.84
16.01
Claude Sonnet 4.6 (Non-reasoning, Low Effort)
1M
AnthropicAnthropic
34*
--
44
1.23
12.54
GLM-5.2
1M
Z AIZ AI
34
--
162
1.43
4.51
GPT-5.6 Terra (Non-reasoning)
1M
OpenAIOpenAI
34
$0.18
123
0.79
4.85
KAT-Coder-Pro V2
256k
KwaiKATKwaiKAT
34
--
103
1.57
6.42
Qwen3.5 397B A17B
262k
AlibabaAlibaba
34
$0.33
69
2.36
55.49
Hy3-preview
256k
TencentTencent
34*
--
152
3.57
20.00
LongCat 2.0
1M
LongCatLongCat
33
--
--
--
--
GPT-5.6 Luna (low)
1M
OpenAIOpenAI
33
$0.06
168
1.53
4.52
MiMo-V2-Flash (Feb 2026)
256k
XiaomiXiaomi
33*
--
--
--
--
Qwen3.5 122B A10B
262k
AlibabaAlibaba
32
$0.24
133
2.32
21.09
Qwen3.5 397B A17B
262k
AlibabaAlibaba
32*
--
69
2.36
9.62
Qwen3.6 35B A3B
262k
AlibabaAlibaba
32
$0.18
158
2.21
39.49
DeepSeek V4 Pro
1M
DeepSeekDeepSeek
31*
--
73
1.52
8.32
Qwen3.5 Omni Plus
256k
AlibabaAlibaba
31*
--
53
2.37
11.77
Ring-2.6-1T
262k
InclusionAIInclusionAI
31
$0.35
126
3.37
23.28
Qwen3.6 27B
262k
AlibabaAlibaba
30
$0.36
57
3.63
12.40
o3
200k
OpenAIOpenAI
30*
--
139
6.60
10.20
Step 3.7 Flash
262k
StepFunStepFun
30
$0.09
396
0.89
7.20
Mistral Medium 3.5
256k
MistralMistral
30
$0.39
104
1.91
26.03
Claude 4.5 Haiku
200k
AnthropicAnthropic
30
$0.24
104
13.30
18.08
Gemma 4 31B
256k
GoogleGoogle
29
--
35
1.12
65.03
GPT-5.5 Instant (June 2026)
400k
OpenAIOpenAI
29
$0.54
--
--
--
DeepSeek V4 Flash
1M
DeepSeekDeepSeek
29*
--
117
1.22
5.49
JT-35B-Flash
256k
China MobileChina Mobile
28*
--
--
--
--
KAT-Coder-Pro V1
256k
KwaiKATKwaiKAT
28*
--
--
--
--
MiMo-V2.5-Pro
1M
XiaomiXiaomi
28*
--
67
2.77
10.24
Qwen3.5 122B A10B
262k
AlibabaAlibaba
28
$0.18
146
2.37
5.80
GPT-5.6 Luna (Non-reasoning)
1M
OpenAIOpenAI
27
$0.08
179
0.79
3.59
Hy3-preview
256k
TencentTencent
26*
--
150
3.70
7.03
Ling-2.6-1T
262k
InclusionAIInclusionAI
26*
--
--
--
--
Step 3.5 Flash 2603
256k
StepFunStepFun
26*
--
274
1.06
10.17
Doubao Seed Code
256k
ByteDance SeedByteDance Seed
26*
--
--
--
--
Gemini 2.5 Pro
1M
GoogleGoogle
26
$0.20
139
23.89
27.48
Gemma 4 26B A4B
256k
GoogleGoogle
26
$0.03
--
--
--
NVIDIA Nemotron 3 Super
1M
NVIDIANVIDIA
25
$0.21
187
1.65
15.00
Gemini 3.1 Flash-Lite
1M
GoogleGoogle
25
$0.04
314
6.01
7.60
Grok 4.3 (Non-reasoning)
1M
SpaceXAISpaceXAI
25
$0.37
100
0.84
5.84
MiMo-V2-Flash
256k
XiaomiXiaomi
25
--
--
--
--
Qwen3.6 35B A3B
262k
AlibabaAlibaba
24
$0.60
190
2.26
4.89
Qwen3.5 35B A3B
262k
AlibabaAlibaba
24
$0.23
134
2.17
5.90
gpt-oss-120b (high)
131k
OpenAIOpenAI
24
$0.06
288
0.86
9.54
Claude 4.5 Haiku
200k
AnthropicAnthropic
24*
--
93
0.89
6.28
Command A+
192k
CohereCohere
23
$0.00
203
0.40
12.70
K-EXAONE
256k
LG AI ResearchLG AI Research
22
--
--
--
--
ERNIE 5.0 Thinking Preview
128k
BaiduBaidu
22*
--
--
--
--
Gemma 4 12B
256k
GoogleGoogle
22
--
110
2.43
25.18
Gemma 4 31B
256k
GoogleGoogle
22
$0.03
73
2.24
9.06
Nova 2.0 Pro Preview (medium)
256k
AmazonAmazon
22
$0.17
117
15.71
37.16
Qwen3.5 9B
262k
AlibabaAlibaba
21
$0.22
71
1.57
36.67
Mercury 2
128k
InceptionInception
21
$0.08
987
4.00
4.51
Qwen3 Coder Next
256k
AlibabaAlibaba
21
$0.33
128
1.25
5.15
Nova 2.0 Omni (medium)
1M
AmazonAmazon
21*
--
--
--
--
Apriel-v1.6-15B-Thinker
128k
ServiceNowServiceNow
21*
--
--
--
--
Qwen3.5 9B
262k
AlibabaAlibaba
20*
--
--
--
--
EXAONE 4.5 33B
262k
LG AI ResearchLG AI Research
20
--
--
--
--
Gemma 4 26B A4B
256k
GoogleGoogle
20*
--
75
1.16
7.83
Qwen3.5 4B
262k
AlibabaAlibaba
20*
--
31
0.76
82.08
North Mini Code
256k
CohereCohere
20
$0.00
64
0.48
39.45
Nova 2.0 Pro Preview (low)
256k
AmazonAmazon
20
$0.21
120
9.70
30.55
Mistral Small 4
256k
MistralMistral
20
$0.10
163
0.75
16.08
Devstral 2
256k
MistralMistral
19
$0.00
29
1.27
18.35
Nova 2.0 Lite (medium)
1M
AmazonAmazon
19*
--
151
21.47
38.01
Qwen3.5 Omni Flash
256k
AlibabaAlibaba
19*
--
253
1.87
3.85
JT-MINI
128k
China MobileChina Mobile
19*
--
--
--
--
Nova 2.0 Lite (high)
1M
AmazonAmazon
18
$0.25
149
18.92
35.71
Trinity Large Thinking
512k
Arcee AIArcee AI
18
$0.13
172
0.90
15.44
Magistral Medium 1.2
128k
MistralMistral
18
$0.75
44
1.71
58.62
Nova 2.0 Lite (low)
1M
AmazonAmazon
18*
--
150
9.80
26.46
HyperNova 60B 2605
131k
Multiverse ComputingMultiverse Computing
18
$0.02
367
0.89
7.69
Nemotron Cascade 2 30B A3B
1M
NVIDIANVIDIA
18
--
--
--
--
Devstral Small 2
256k
MistralMistral
17
$0.00
30
1.35
18.30
K2 Think V2
262k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
17
--
--
--
--
LongCat Flash Lite
256k
LongCatLongCat
17*
--
--
--
--
HyperCLOVA X SEED Think (32B)
128k
NaverNaver
17*
--
--
--
--
K-EXAONE
256k
LG AI ResearchLG AI Research
17*
--
--
--
--
Qwen3 Next 80B A3B
262k
AlibabaAlibaba
17
$0.17
208
2.24
14.29
Nova 2.0 Omni (low)
1M
AmazonAmazon
17*
--
--
--
--
Mi:dm K 2.5 Pro
128k
Korea TelecomKorea Telecom
16*
--
--
--
--
G9v3-3B
131k
AI9StarsAI9Stars
16
--
--
--
--
Qwen3.5 4B
262k
AlibabaAlibaba
16*
--
30
0.71
17.61
Mistral Large 3
256k
MistralMistral
16
$0.06
49
1.14
11.28
INTELLECT-3
131k
Prime IntellectPrime Intellect
16*
--
--
--
--
Solar Open 100B
128k
UpstageUpstage
15*
--
--
--
--
Nemotron 3 Nano Omni 30B A3B Reasoning
256k
NVIDIANVIDIA
15*
--
319
0.95
8.77
gpt-oss-120b (low)
131k
OpenAIOpenAI
15
$0.02
293
0.84
9.36
gpt-oss-20b (high)
131k
OpenAIOpenAI
15
$0.02
197
0.83
13.55
Nova 2.0 Pro Preview
256k
AmazonAmazon
14
$0.25
106
1.01
5.72
gpt-oss-20b (low)
131k
OpenAIOpenAI
14*
--
222
0.81
12.06
Llama 4 Maverick
1M
MetaMeta
14
$0.03
105
0.92
5.67
K2-V2 (high)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
14*
--
--
--
--
NVIDIA Nemotron 3 Nano
1M
NVIDIANVIDIA
14
$0.02
157
1.52
17.49
Solar Pro 3
128k
UpstageUpstage
14
--
--
--
--
Ling 2.6 Flash
262k
InclusionAIInclusionAI
14
--
149
1.12
4.47
Qwen3 Next 80B A3B
262k
AlibabaAlibaba
14*
--
180
2.27
5.05
Tri-21B-think Preview
32k
Trillion LabsTrillion Labs
14*
--
--
--
--
DiffusionGemma 26B A4B
256k
GoogleGoogle
13
--
--
--
--
Gemma 4 12B (Non-reasoning)
262k
GoogleGoogle
13*
--
105
2.39
7.14
Motif-2-12.7B
128k
Motif TechnologiesMotif Technologies
13*
--
--
--
--
Nova Premier
1M
AmazonAmazon
13*
--
33
2.84
18.03
K2-V2 (medium)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
12*
--
--
--
--
Llama Nemotron Super 49B v1.5
128k
NVIDIANVIDIA
12*
--
83
10.29
40.55
Mistral Small 4
256k
MistralMistral
12*
--
151
0.68
4.00
Tri-21B-Think
32k
Trillion LabsTrillion Labs
12*
--
--
--
--
MiniCPM5-1B
128k
OpenBMBOpenBMB
12*
--
--
--
--
Sarvam 105B (high)
128k
SarvamSarvam
12*
--
--
--
--
Gemma 4 E4B
128k
GoogleGoogle
12
--
92
0.78
28.09
Nova 2.0 Lite
1M
AmazonAmazon
12*
--
140
1.15
4.73
MiniCPM5-1B
128k
OpenBMBOpenBMB
12*
--
--
--
--
Magistral Small 1.2
128k
MistralMistral
11
$0.25
89
0.93
29.09
Nanbeige4.1-3B
256k
NanbeigeNanbeige
11
--
--
--
--
Ministral 3 14B
256k
MistralMistral
11
$0.15
64
0.85
8.70
EXAONE 4.0 32B
131k
LG AI ResearchLG AI Research
11*
--
--
--
--
Nova 2.0 Omni
1M
AmazonAmazon
11*
--
--
--
--
Llama 4 Scout
10M
MetaMeta
10
$0.01
96
0.76
5.97
Hermes 4 70B
128k
Nous ResearchNous Research
10*
--
93
1.35
28.24
Falcon-H1R-7B
256k
TII UAETII UAE
10*
--
--
--
--
Qwen3 Omni 30B A3B
65.5k
AlibabaAlibaba
10*
--
99
1.92
27.13
Gemma 4 E2B
128k
GoogleGoogle
9
--
--
--
--
Step3 VL 10B
65.5k
StepFunStepFun
9*
--
--
--
--
Llama 3.3 70B
128k
MetaMeta
9
$0.08
85
1.65
7.53
Llama Nemotron Ultra
128k
NVIDIANVIDIA
9*
--
53
2.32
49.56
ERNIE 4.5 300B A47B
131k
BaiduBaidu
9*
--
--
--
--
Hermes 4 405B
128k
Nous ResearchNous Research
9*
--
40
2.36
65.40
Solar Pro 2
65.5k
UpstageUpstage
9*
--
--
--
--
NVIDIA Nemotron Nano 12B v2 VL
128k
NVIDIANVIDIA
9*
--
72
5.57
40.05
Ministral 3 8B
256k
MistralMistral
9
$0.18
115
0.73
5.07
Gemma 4 E4B
128k
GoogleGoogle
9*
--
94
0.75
6.08
Granite 4.1 30B
131k
IBMIBM
9
--
--
--
--
NVIDIA Nemotron Nano 9B V2
131k
NVIDIANVIDIA
9*
--
80
7.00
38.30
Hermes 4 405B
128k
Nous ResearchNous Research
9*
--
40
2.37
14.78
NVIDIA Nemotron 3 Nano 4B
262k
NVIDIANVIDIA
9
--
--
--
--
Llama Nemotron Super 49B v1.5
128k
NVIDIANVIDIA
9*
--
83
4.95
10.96
K2-V2 (low)
512k
MBZUAI Institute of Foundation ModelsMBZUAI Institute of Foundation Models
9*
--
--
--
--
Kimi Linear 48B A3B Instruct
1M
KimiKimi
9*
--
--
--
--
Llama 3.1 405B
128k
MetaMeta
9*
--
--
--
--
LFM2.5-8B-A1B
32.8k
Liquid AILiquid AI
8*
--
335
11.42
18.89
Ring-flash-2.0
128k
InclusionAIInclusionAI
8*
--
--
--
--
Olmo 3.1 32B Think
65.5k
Allen Institute for AIAllen Institute for AI
8*
--
--
--
--
Solar Pro 2
65.5k
UpstageUpstage
8*
--
--
--
--
Command A
256k
CohereCohere
8*
--
57
1.67
10.50
Llama 3.1 Nemotron 70B
128k
NVIDIANVIDIA
8*
--
74
9.88
16.66
NVIDIA Nemotron 3 Nano
1M
NVIDIANVIDIA
7*
--
100
0.93
5.92
NVIDIA Nemotron Nano 9B V2
131k
NVIDIANVIDIA
7*
--
156
1.91
5.12
Hermes 4 70B
128k
Nous ResearchNous Research
7*
--
93
1.34
6.71
Qwen3.5 2B
262k
AlibabaAlibaba
7
--
--
--
--
Granite 4.1 8B
131k
IBMIBM
7*
--
105
0.81
5.58
Sarvam 30B (high)
65.5k
SarvamSarvam
7*
--
--
--
--
Olmo 3.1 32B Instruct
65.5k
Allen Institute for AIAllen Institute for AI
6*
--
--
--
--
Ministral 3 3B
256k
MistralMistral
6
$0.13
247
0.63
2.65
Gemma 4 E2B
128k
GoogleGoogle
6*
--
--
--
--
R1 1776
128k
PerplexityPerplexity
6*
--
--
--
--
Llama 3.2 90B (Vision)
128k
MetaMeta
6*
--
--
--
--
Phi-4 Mini
128k
MicrosoftMicrosoft
6
$0.00
45
0.82
12.04
EXAONE 4.0 32B
131k
LG AI ResearchLG AI Research
6*
--
--
--
--
Qwen3.5 2B
262k
AlibabaAlibaba
6
--
--
--
--
Qwen3.5 0.8B
262k
AlibabaAlibaba
5
--
--
--
--
DeepHermes 3 - Mistral 24B
32k
Nous ResearchNous Research
5*
--
--
--
--
Jamba 1.7 Large
256k
AI21 LabsAI21 Labs
5*
--
54
1.39
10.71
Granite 4.0 H Small
128k
IBMIBM
5*
--
386
10.22
11.52
Qwen3 Omni 30B A3B
65.5k
AlibabaAlibaba
5*
--
94
1.98
7.28
LFM2 24B A2B
32.8k
Liquid AILiquid AI
5*
--
--
--
--
Phi-4
16k
MicrosoftMicrosoft
5*
--
29
4.01
21.17
Nova Micro
130k
AmazonAmazon
5*
--
285
0.91
2.67
Granite 4.1 3B
131k
IBMIBM
5
--
--
--
--
NVIDIA Nemotron Nano 12B v2 VL
128k
NVIDIANVIDIA
5*
--
154
2.48
5.73
Phi-4 Multimodal
128k
MicrosoftMicrosoft
5*
--
18
0.81
28.98
MiniCPM-V 4.6 1.3B
262k
OpenBMBOpenBMB
4
--
--
--
--
Jamba Reasoning 3B
262k
AI21 LabsAI21 Labs
4*
--
--
--
--
Reka Flash 3
128k
Reka AIReka AI
4*
--
--
--
--
Olmo 3 7B Think
65.5k
Allen Institute for AIAllen Institute for AI
4*
--
--
--
--
Molmo 7B-D
4.1k
Allen Institute for AIAllen Institute for AI
4*
--
--
--
--
Ling-mini-2.0
131k
InclusionAIInclusionAI
4*
--
--
--
--
Llama 3.2 11B (Vision)
128k
MetaMeta
3*
--
6
2.65
81.68
Qwen3.5 0.8B
262k
AlibabaAlibaba
3
--
--
--
--
Exaone 4.0 1.2B
64k
LG AI ResearchLG AI Research
3*
--
--
--
--
Olmo 3 7B
65.5k
Allen Institute for AIAllen Institute for AI
3*
--
--
--
--
Exaone 4.0 1.2B
64k
LG AI ResearchLG AI Research
3*
--
--
--
--
LFM2.5-1.2B-Thinking
32k
Liquid AILiquid AI
3*
--
--
--
--
Jamba 1.7 Mini
258k
AI21 LabsAI21 Labs
3*
--
--
--
--
LFM2 2.6B
32.8k
Liquid AILiquid AI
3*
--
--
--
--
LFM2.5-1.2B-Instruct
32k
Liquid AILiquid AI
3*
--
--
--
--
Granite 4.0 H 1B
128k
IBMIBM
3*
--
--
--
--
Gemma 3 270M
32k
GoogleGoogle
2*
--
--
--
--
Apertus 70B Instruct
65.5k
Swiss AI InitiativeSwiss AI Initiative
2*
--
--
--
--
Granite 4.0 Micro
128k
IBMIBM
2*
--
--
--
--
DeepHermes 3 - Llama-3.1 8B
128k
Nous ResearchNous Research
2*
--
--
--
--
Granite 4.0 1B
128k
IBMIBM
2*
--
--
--
--
Molmo2-8B
36.9k
Allen Institute for AIAllen Institute for AI
2*
--
--
--
--
LFM2 8B A1B
32.8k
Liquid AILiquid AI
2*
--
--
--
--
LFM2.5-VL-1.6B
32k
Liquid AILiquid AI
1*
--
400
11.48
12.73
Granite 4.0 350M
32.8k
IBMIBM
1*
--
--
--
--
Tiny Aya Global
8.19k
CohereCohere
1*
--
--
--
--
Apertus 8B Instruct
65.5k
Swiss AI InitiativeSwiss AI Initiative
1*
--
--
--
--
Granite 4.0 H 350M
32.8k
IBMIBM
1*
--
--
--
--
Claude Sonnet 5 (low)
1M
AnthropicAnthropic
--
--
60
2.06
10.44
EXAONE 4.5 33B
262k
LG AI ResearchLG AI Research
--
--
--
--
--
Claude Sonnet 5 (xhigh)
1M
AnthropicAnthropic
--
--
68
28.07
35.45
Gemini 3 Deep Think
128k
GoogleGoogle
--
--
--
--
--
Claude Sonnet 5 (high)
1M
AnthropicAnthropic
--
--
63
7.95
15.87
Mi:dm K 2.5 Pro Preview
128k
Korea TelecomKorea Telecom
--
--
--
--
--
GPT-5.5 Pro (xhigh)
922k
OpenAIOpenAI
--
--
--
--
--
Cogito v2.1
128k
Deep CogitoDeep Cogito
--
--
--
--
--
Claude Sonnet 5 (medium)
1M
AnthropicAnthropic
--
--
63
2.26
10.20

Definiciones clave

Maximum number of combined input & output tokens. Output tokens commonly have a significantly lower limit (varied by model).

Tokens per second received while the model is generating tokens (ie. after first chunk has been received from the API for models which support streaming).

Time to first token received, in seconds, after API request sent. For reasoning models which share reasoning tokens, this will be the first reasoning token. For models which do not support streaming, this represents time to receive the completion.

Average cost per task in the index. Costs are split by input, cache hit, cache write, reasoning, and answer token pricing where canonical token counts are available.

Price per token included in the request/message sent to the API, represented as USD per million Tokens.

Price per token for cached prompts (previously processed), typically offering a significant discount compared to regular input price, represented as USD per million tokens. The values shown here are the cache hit price; cache write and cache storage are billed separately and vary by provider — see "Cache pricing by provider" for detail.

Price per token to write prompt tokens into the cache so that later requests can hit them, represented as USD per million tokens. Some providers charge a premium over the standard input price to create a cache entry (e.g. Anthropic), while others cache automatically with no separate write fee.

Price per token generated by the model (received from the API), represented as USD per million Tokens.

Metrics are 'live' and are based on the past 72 hours of measurements, measurements are taken 8 times a day for single requests and 2 times per day for parallel requests.

Preguntas frecuentes

Claude Opus 5 (Adaptive Reasoning, Max Effort) ocupa actualmente el puesto #1 en el ranking de modelos LLM de Artificial Analysis, con una puntuación de 61 en el Índice de Inteligencia, entre 170 modelos clasificados.

Los mejores modelos por Índice de Inteligencia son: 1. Claude Opus 5 (Adaptive Reasoning, Max Effort) (61), 2. Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) (60), 3. Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) (60), 4. GPT-5.6 Sol (max) (59), 5. Claude Opus 5 (Adaptive Reasoning, High Effort) (59).

Mercury 2 es el más rápido con 986.7 tokens por segundo, seguido por Gemini 3.5 Flash-Lite (432.7 t/s) y LFM2.5-VL-1.6B (400.0 t/s).

Llama 4 Scout tiene el menor costo por tarea del Índice de Inteligencia con $0.01, seguido por MiMo-V2.5 ($0.01) y gpt-oss-120b (low) ($0.02).

Kimi K3 es el modelo con pesos abiertos mejor clasificado, con una puntuación de 57 en el Índice de Inteligencia. Hay 95 modelos con pesos abiertos de un total de 170 en el ranking.

Los mejores modelos con pesos abiertos por Índice de Inteligencia son: 1. Kimi K3 (57), 2. GLM-5.2 (max) (51), 3. MiniMax-M3 (44).

Claude Opus 5 (Adaptive Reasoning, Max Effort) lidera entre 126 modelos de razonamiento, con una puntuación de 61 en el Índice de Inteligencia. Los modelos de razonamiento usan pensamiento extendido para resolver problemas complejos antes de responder.

El ranking incluye filtros para limitar los resultados por tipo de modelo (razonamiento o no razonamiento), apertura (pesos abiertos o propietario) y otros criterios. También puedes ajustar las opciones de prompt para ver cómo varía el rendimiento con distintas longitudes de entrada.

Haz clic en cualquier nombre de modelo en la clasificación para visitar su página de comparación dedicada, con gráficos detallados sobre inteligencia, precios, velocidad, latencia y más. También puedes comparar proveedores de API para cada modelo. Ver todos los modelos