CoreWeave: inteligencia, rendimiento y precio de sus modelos

CoreWeave
CoreWeave

Este análisis está pensado para ayudarte a elegir el mejor modelo ofrecido por CoreWeave para tu caso de uso.

Más inteligente

Updated
#1
GLM-5.3-FlashGLM-5.3-Flash
42
#2
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
39
#3
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
34
#4
GLM-5.2 (max)GLM-5.2 (max)
34
#5
Qwen3.8 27B (xhigh) (FP8)Qwen3.8 27B (xhigh) (FP8)
34

Índice de inteligencia

27 modelos en total

Más rápido

#1
Nemotron 3 UltraNemotron 3 Ultra
298 t/s
#2
Kimi K2.7 CodeKimi K2.7 Code
289 t/s
#3
Kimi K2.6Kimi K2.6
213 t/s
#4
Qwen3.6 35B A3B (Non-reasoning) (FP8)Qwen3.6 35B A3B (Non-reasoning) (FP8)
188 t/s
#5
gpt-oss-20b (high)gpt-oss-20b (high)
183 t/s

Velocidad de salida

27 modelos en total

Menor precio

#1
gpt-oss-20b (low)gpt-oss-20b (low)
$0.04
#2
gpt-oss-20b (high)gpt-oss-20b (high)
$0.04
#3
gpt-oss-120b (high)gpt-oss-120b (high)
$0.04
#4
gpt-oss-120b (low)gpt-oss-120b (low)
$0.04
#5
Granite 4.2 8BGranite 4.2 8B
$0.07

Precio combinado (por 1M de tokens)

27 modelos en total

CoreWeave ofrece 27 modelos, cada uno con distintas características de inteligencia, rendimiento y precio. A continuación se comparan las métricas clave entre modelos.

  • En inteligencia, los mejores modelos en CoreWeave son GLM-5.3-Flash (42), DeepSeek V4.1 Flash (max) (39) y DeepSeek V4 Flash 0731 (max) (34).
  • En velocidad de salida, los modelos más rápidos son Nemotron 3 Ultra (298 t/s), Kimi K2.7 Code (289 t/s) y Kimi K2.6 (213 t/s). La velocidad varía significativamente entre modelos, con una diferencia del 62% entre el más rápido y el más lento.
  • En latencia, Gemma 4 26B A4B (Non-reasoning) (0.59s), Llama 3.1 8B (0.70s) y Llama 3.3 70B (0.90s) ofrecen el menor tiempo hasta el primer token de respuesta.
  • En precios, gpt-oss-20b (low) ($0.04), gpt-oss-20b (high) ($0.04) y gpt-oss-120b (high) ($0.04) ofrecen los menores precios combinados por 1M de tokens.
  • En tamaño de ventana de contexto, GLM-5.3-Flash (1M), DeepSeek V4.1 Flash (max) (1M) y DeepSeek V4 Pro (max) (1M) admiten las ventanas de contexto más grandes en CoreWeave.

Aspectos destacados

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Evaluaciones de inteligencia

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Agentic tool use

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Ventana de contexto

Context Window

Context window: tokens limit · Higher is better

Precios

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Resumen de rendimiento

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Velocidad

Medida por la velocidad de salida (tokens por segundo)

Output Speed

Output tokens per second · Higher is better

Latencia

Medida por el tiempo (segundos) hasta el primer token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

Tiempo de respuesta de extremo a extremo

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Análisis adicional
Logo de Z AI
GLM-5.3-Flash
1M
Abierto
42
$0.76
171
1.65
16.29
11.71
Logo de DeepSeek
DeepSeek V4.1 Flash (max)
1M
Abierto
39
$0.62
186
0.92
14.33
10.73
Logo de DeepSeek
DeepSeek V4 Flash 0731 (max)
262k
Abierto
34
$0.44
123
1.69
22.07
16.31
Logo de Z AI
GLM-5.2 (max)
262k
Abierto
34
$0.70
68
1.42
38.00
29.27
Logo de Alibaba
Qwen3.8 27B (xhigh) (FP8)
262k
Abierto
34
$0.92
66
1.72
39.61
30.31
Logo de DeepSeek
DeepSeek V4 Pro (max)
1M
Abierto
30
$2.00
178
1.29
28.69
24.60
Logo de MiniMax
MiniMax-M3 (NVFP4)
262k
Abierto
29
--
69
1.73
38.06
29.06
Logo de Z AI
GLM-5 (FP8)
200k
Abierto
28*
--
--
--
--
--
Logo de Kimi
Kimi K2.6
262k
Abierto
27
$1.18
212
1.25
24.64
21.03
Logo de Kimi
Kimi K2.7 Code
262k
Abierto
26
$0.83
299
1.10
10.22
7.45
Logo de Kimi
Kimi K2.6 (Non-reasoning)
262k
Abierto
24*
--
126
1.15
5.13
--
Logo de NVIDIA
Nemotron 3 Ultra
262k
Abierto
23
$0.51
313
0.99
9.85
7.26
Logo de DeepSeek
DeepSeek V4 Pro (Non-reasoning)
1M
Abierto
21*
--
161
1.30
4.42
--
Logo de Google
Gemma 4 31B
262k
Abierto
19*
--
61
2.14
38.84
28.50
Logo de Alibaba
Qwen3.6 35B A3B (FP8)
262k
Abierto
18
$0.31
178
1.05
34.20
30.34
Logo de Google
Gemma 4 26B A4B
262k
Abierto
17*
--
94
0.65
27.12
21.17
Logo de Alibaba
Qwen3.6 35B A3B (Non-reasoning) (FP8)
262k
Abierto
15*
--
184
1.06
3.78
--
Logo de Google
Gemma 4 31B (Non-reasoning)
262k
Abierto
14*
--
38
2.11
15.25
--
Logo de DeepSeek
DeepSeek V3.1 (Non-reasoning)
128k
Abierto
14*
--
69
1.40
8.61
--
Logo de Google
Gemma 4 26B A4B (Non-reasoning)
262k
Abierto
13*
--
94
0.60
5.93
--
Logo de NVIDIA
Nemotron 3.5 Lightning (BF16)
262k
Abierto
13
$0.10
159
0.64
16.32
12.54
Logo de OpenAI
gpt-oss-120b (high)
131k
Abierto
12
$0.02
56
1.35
46.21
35.89
Logo de IBM
Granite 4.2 8B
131k
Abierto
11
$0.05
68
0.77
37.71
29.55
Logo de OpenAI
gpt-oss-120b (low)
131k
Abierto
10*
--
52
1.39
49.16
38.21
Logo de OpenAI
gpt-oss-20b (low)
131k
Abierto
10*
--
157
0.61
16.49
12.71
Logo de OpenAI
gpt-oss-20b (high)
131k
Abierto
9
$0.01
183
0.62
14.26
10.91
Logo de Meta
Llama 3.3 70B
128k
Abierto
8*
--
84
0.91
6.90
--
Logo de Meta
Llama 3.1 8B
128k
Abierto
7*
--
140
0.71
4.29
--

Definiciones clave

Preguntas frecuentes

Preguntas comunes sobre CoreWeave