Todos os lançamentos

Ferramenta de comparação de lançamentos

Compare até cinco lançamentos de modelos lado a lado, com todas as variantes de cada um, em inteligência, preços, velocidade de saída, latência, janela de contexto e mais.

Para saber mais sobre como medimos e avaliamos os modelos, consulte a página de metodologia.

Playground de modelos

Lançamentos comparados

Claude Opus 5.5 (max with fallback)Claude Opus 5.5 (xhigh with fallback)Claude Opus 5.5 (high with fallback)GPT-6 Astra (max)GPT-6 Astra (xhigh)Claude Opus 5.5 (medium with fallback)GPT-6 Astra (high)GPT-6 Astra (medium)GPT-6 Astra (low)Claude Opus 5.5 (low with fallback)
Intelligence Index
58
56
54
53
52
51
51
50
46
42
Finanças e Contabilidade
61
58
56
55
54
54
53
52
49
46
Estratégia e Operações
64
62
59
57
57
57
55
54
50
48
Jurídico
63
61
59
59
58
57
58
57
53
52
Saúde e Medicina
61
--
--
52
--
--
--
--
--
--
Custo por tarefaUSD
$5.98
$3.46
$1.82
$3.26
$2.31
$1.34
$1.73
$1.54
$0.82
$0.55
Velocidade de saídaTokens/s
92
79
72
59
55
73
53
51
51
71
Tempo até o primeiro tokens
710,36
147,01
34,40
276,50
142,98
26,57
67,00
4,50
2,65
4,50
Janela de contexto
1M
1M
1M
1M
1M
1M
1M
1M
1M
1M
Criador
Anthropic
Anthropic
Anthropic
OpenAI
OpenAI
Anthropic
OpenAI
OpenAI
OpenAI
Anthropic
Licença
Proprietário
Proprietário
Proprietário
Proprietário
Proprietário
Proprietário
Proprietário
Proprietário
Proprietário
Proprietário
Modalidade de entrada

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Compatível com: texto e imagem

Modalidade de saída

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Compatível com: texto

Provedores
Anthropic
Anthropic
Anthropic
Microsoft AzureOpenAIAmazon Bedrock
OpenAI
Anthropic
OpenAI
OpenAI
Microsoft AzureOpenAI
AnthropicAmazon BedrockGoogle
Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
Weighted average cost (USD) per Intelligence Index task · Lower is better

Inteligência

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Not publicly available

Intelligence Index vs. custo por tarefa do Intelligence Index

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Opus 5.5
GPT-6 Astra
Pareto line

Custo por tarefa (USD, escala logarítmica)

Custo

Custo por tarefa do Intelligence Index

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better
Not publicly available

Velocidade e latência

Velocidade de saída

Output tokens per second · Higher is better

Pontuações de capacidades

Índices de capacidades

Mede o desempenho dos modelos em capacidades e setores específicos
Índice de Finanças e Contabilidade

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Índice de Estratégia e Operações

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Índice Jurídico

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Índice de Saúde e Medicina

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Índice de Engenharia

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

Índice de Economia

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better