Claude Opus 5 vs. Mistral Small 4: comparación de lanzamientos

Comparación de los lanzamientos Claude Opus 5 y Mistral Small 4: 5 y 2 modelos respectivamente, cada uno con distintas características de inteligencia, rendimiento y precio. A continuación se comparan las métricas clave de los 7 modelos.

  • En inteligencia, Claude Opus 5 obtiene la puntuación más alta: Claude Opus 5 (Adaptive Reasoning, Max Effort), con 51, frente a Mistral Small 4 (Reasoning), con 11 de Mistral Small 4.
  • En velocidad de salida, Mistral Small 4 es el más rápido: Mistral Small 4 (Reasoning), con 159 tokens/s, frente a Claude Opus 5 (Adaptive Reasoning, Medium Effort), con 57 tokens/s de Claude Opus 5.
  • En costo por tarea, Mistral Small 4 es el más económico: Mistral Small 4 (Reasoning), con $0.05, frente a Claude Opus 5 (Adaptive Reasoning, Low Effort), con $1.10 de Claude Opus 5.

Lanzamientos comparados

Inteligencia

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Opus 5
Pareto line

Cost per Task (USD, Log Scale)

Costo

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Velocidad y latencia

Output Speed

Output tokens per second · Higher is better

Puntuaciones de capacidades

Índices de capacidades

Mide el rendimiento de los modelos en capacidades e industrias específicas
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better

Más detalles

Modalidad de entrada
Modalidad de salida
Pesos
Benchmarks de proveedores
Claude Opus 5 (Adaptive Reasoning, Max Effort)
Logo de AnthropicAnthropic
51
55
57
57
53
54
61
$5.86
3,9 US$
$5.00
$25.00
$0.50
7275 US$
73k
43k
140M
55
62.61s
62.61s
71.64s
803.97s
-
1M
jul 2026
-

Admite: texto e imágenes

Admite: texto

No disponible
-
Amazon BedrockAnthropicGoogle
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
Logo de AnthropicAnthropic
50
54
54
56
53
53
60
$4.88
3,9 US$
$5.00
$25.00
$0.50
5868 US$
61k
35k
111M
54
26.58s
26.58s
35.86s
700.83s
-
1M
jul 2026
-

Admite: texto e imágenes

Admite: texto

No disponible
-
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, High Effort)
Logo de AnthropicAnthropic
48
51
53
54
52
52
58
$3.61
3,9 US$
$5.00
$25.00
$0.50
4332 US$
46k
25k
81M
56
16.08s
16.08s
25.04s
508.17s
-
1M
jul 2026
-

Admite: texto e imágenes

Admite: texto

No disponible
-
AnthropicGoogleAmazon Bedrock
Claude Opus 5 (Adaptive Reasoning, Medium Effort)
Logo de AnthropicAnthropic
45
49
51
52
50
47
56
$2.19
3,9 US$
$5.00
$25.00
$0.50
2732 US$
29k
15k
49M
57
5.20s
5.20s
13.99s
308.27s
-
1M
jul 2026
-

Admite: texto e imágenes

Admite: texto

No disponible
-
AnthropicAmazon BedrockGoogle
Claude Opus 5 (Adaptive Reasoning, Low Effort)
Logo de AnthropicAnthropic
39
44
46
48
45
42
50
$1.10
3,9 US$
$5.00
$25.00
$0.50
1561 US$
15k
7k
26M
57
3.04s
3.04s
11.84s
156.93s
-
1M
jul 2026
-

Admite: texto e imágenes

Admite: texto

No disponible
-
AnthropicAmazon BedrockGoogle
Mistral Small 4 (Reasoning)
Logo de MistralMistral
11
10
7
13
-
16
18
$0.05
0,2 US$
$0.15
$0.60
-
79 US$
14k
9k
54M
159
0.74s
13.28s
16.42s
93.03s
119B
6,5B activos en inferencia
256k
mar 2026
-

Admite: texto e imágenes

Admite: texto

Mistral
Mistral Small 4 (Non-reasoning)
Logo de MistralMistral
9
-
-
-
-
-
-
-
0,2 US$
$0.15
$0.60
-
-
-
-
-
153
0.72s
0.72s
3.98s
-
119B
6,5B activos en inferencia
256k
mar 2026
-
No

Admite: texto e imágenes

Admite: texto

Mistral