Makora: inteligencia, rendimiento y precio de sus modelos

Makora
Makora

Este análisis está pensado para ayudarte a elegir el mejor modelo ofrecido por Makora para tu caso de uso.

Más inteligente

Updated
#1
GLM-5.3 (max)GLM-5.3 (max)
45
#2
Kimi K3 (max)Kimi K3 (max)
44
#3
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
39
#4
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
34
#5
DeepSeek V4 Flash (high)DeepSeek V4 Flash (high)
26

Índice de inteligencia

8 modelos en total

Más rápido

#1
Gemma 4 26B A4B (Non-reasoning)Gemma 4 26B A4B (Non-reasoning)
365 t/s
#2
Gemma 4 26B A4BGemma 4 26B A4B
341 t/s
#3
GLM-5.3 (max)GLM-5.3 (max)
315 t/s
#4
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
315 t/s
#5
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
106 t/s

Velocidad de salida

8 modelos en total

Menor precio

#1
Gemma 4 26B A4BGemma 4 26B A4B
$0.08
#2
Gemma 4 26B A4B (Non-reasoning)Gemma 4 26B A4B (Non-reasoning)
$0.08
#3
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
$0.10
#4
DeepSeek V4 Flash (high)DeepSeek V4 Flash (high)
$0.10
#5
DeepSeek V4 Flash (Non-reasoning)DeepSeek V4 Flash (Non-reasoning)
$0.10

Precio combinado (por 1M de tokens)

8 modelos en total

Makora ofrece 8 modelos, cada uno con distintas características de inteligencia, rendimiento y precio. A continuación se comparan las métricas clave entre modelos.

  • En inteligencia, los mejores modelos en Makora son GLM-5.3 (max) (45), Kimi K3 (max) (44) y DeepSeek V4.1 Flash (max) (39).
  • En velocidad de salida, los modelos más rápidos son Gemma 4 26B A4B (Non-reasoning) (365 t/s), Gemma 4 26B A4B (341 t/s) y GLM-5.3 (max) (315 t/s). La velocidad varía significativamente entre modelos, con una diferencia del 243% entre el más rápido y el más lento.
  • En latencia, DeepSeek V4 Flash (Non-reasoning) (0.73s), Gemma 4 26B A4B (Non-reasoning) (0.82s) y Gemma 4 26B A4B (6.64s) ofrecen el menor tiempo hasta el primer token de respuesta.
  • En precios, Gemma 4 26B A4B ($0.08), Gemma 4 26B A4B (Non-reasoning) ($0.08) y DeepSeek V4 Flash 0731 (max) ($0.10) ofrecen los menores precios combinados por 1M de tokens.
  • En tamaño de ventana de contexto, GLM-5.3 (max) (1M), Kimi K3 (max) (1M) y DeepSeek V4.1 Flash (max) (1M) admiten las ventanas de contexto más grandes en Makora.

Aspectos destacados

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Evaluaciones de inteligencia

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Agentic tool use

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Ventana de contexto

Context Window

Context window: tokens limit · Higher is better

Precios

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Resumen de rendimiento

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant

Velocidad

Medida por la velocidad de salida (tokens por segundo)

Output Speed

Output tokens per second · Higher is better

Latencia

Medida por el tiempo (segundos) hasta el primer token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

Tiempo de respuesta de extremo a extremo

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant

Análisis adicional
Logo de Z AI
GLM-5.3 (max)
1.05M
Abierto
45
$1.90
296
0.85
9.31
6.76
Logo de Kimi
Kimi K3 (max)
1.05M
Abierto
44
$1.21
51
1.20
50.49
39.43
Logo de DeepSeek
DeepSeek V4.1 Flash (max)
1.05M
Abierto
39
$1.11
294
0.64
9.15
6.81
Logo de DeepSeek
DeepSeek V4 Flash 0731 (max)
1M
Abierto
34
$0.42
111
0.84
23.42
18.06
Logo de DeepSeek
DeepSeek V4 Flash (high)
1M
Abierto
26*
--
120
0.82
15.30
10.32
Logo de DeepSeek
DeepSeek V4 Flash (Non-reasoning)
1M
Abierto
19*
--
94
0.73
6.04
--
Logo de Google
Gemma 4 26B A4B
1M
Abierto
17*
--
341
0.77
8.10
5.86
Logo de Google
Gemma 4 26B A4B (Non-reasoning)
1M
Abierto
13*
--
371
0.85
2.20
--

Definiciones clave

Preguntas frecuentes

Preguntas comunes sobre Makora