Idioma Português — Benchmark de modelos de IA Compare o desempenho multilíngue dos LLMs
Os 5 melhores modelos de IA para tarefas em português são Gemini 3.1 Pro Preview, Claude Opus 4.6 (max), Gemini 3 Pro Preview (high), Claude Opus 4.6 (high) e Grok 4. Eles alcançam as maiores pontuações de raciocínio em português no Artificial Analysis Multilingual Index.
Para comparar o desempenho em todos os idiomas disponíveis, consulte a página completa do benchmark multilíngue de modelos de IA.
🇵🇹 Melhores modelos para o idioma Português
#1
Gemini 3.1 Pro Preview94#2
Claude Opus 4.6 (max)94#3
Gemini 3 Pro Preview (high)93#4
Claude Opus 4.6 (high)93#5
Grok 493
Destaques
Multilingual Index
Multilingual Index: idioma Português
Artificial Analysis Multilingual Index · Higher is better
Reasoning models are indicated by a lightbulb icon
Multilingual Index: idioma Português vs. preço
Artificial Analysis Multilingual Index · USD per 1M tokens (blended)
Most attractive quadrant
Reasoning models are indicated by a lightbulb icon
Multilingual Index: idioma Português vs. velocidade de saída
Artificial Analysis Multilingual Index · Output speed: output tokens per second
Most attractive quadrant
Reasoning models are indicated by a lightbulb icon
Multilingual Index: idioma Português vs. janela de contexto
Artificial Analysis Multilingual Index · Context window: tokens limit
Most attractive quadrant
Reasoning models are indicated by a lightbulb icon
Global-MMLU-Lite
Global-MMLU-Lite multilíngue: idioma Português
Multilingual Global-MMLU-Lite · Higher is better
Reasoning models are indicated by a lightbulb icon
Preços
Pricing: Cache Hit, Input, and Output
Price (USD per M Tokens)
Reasoning models are indicated by a lightbulb icon
Velocidade e latência
Output Speed
Output tokens per second · Higher is better
Reasoning models are indicated by a lightbulb icon
Latency: Time To First Answer Token
Seconds to first answer token received · Accounts for reasoning model 'thinking' time
Reasoning models are indicated by a lightbulb icon
End-to-End Response Time
Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better
Reasoning models are indicated by a lightbulb icon