Modelos de DeepSeek: inteligência, desempenho e preço
Artificial Analysis Intelligence Index
A Artificial Analysis avaliou 36 modelos de DeepSeek. Abaixo, comparamos as principais métricas desses modelos.
- Em inteligência, o principal modelo de DeepSeek é DeepSeek V4.1 Flash (Reasoning, Max Effort), com 39.
- Em velocidade de saída, o modelo mais rápido é DeepSeek V4.1 Flash (Non-Reasoning), com 239 tokens/s.
- Em latência, DeepSeek V4.1 Flash (Non-Reasoning), com 1,12 s oferece o menor tempo até o primeiro token de resposta.
- Em termos de preço, DeepSeek V4.1 Flash (Non-Reasoning), com $0.15 oferece o menor custo por tarefa. Os preços variam em até 4,6 vezes entre os modelos.
Inteligência
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Custo
Cost per Intelligence Index Task
Velocidade e latência
Output Speed
Pontuações de capacidades
Índices de capacidades
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Todos os lançamentos de DeepSeek
DeepSeek V4.1 Flash
2 modelos
39
Maior inteligência
DeepSeek V4 Flash Vision
1 modelo
35
Maior inteligência
DeepSeek V4 Pro 0813
1 modelo
36
Maior inteligência
DeepSeek V4 Flash 0731
1 modelo
34
Maior inteligência
DeepSeek V4 Flash 0420
3 modelos
26
Maior inteligência
DeepSeek V4 Pro 0424
3 modelos
30
Maior inteligência
DeepSeek V3.2
2 modelos
21
Maior inteligência
DeepSeek V3.2 Speciale
1 modelo
14
Maior inteligência
DeepSeek V3.2 Exp
2 modelos
17
Maior inteligência
DeepSeek V3.1 Terminus
2 modelos
15
Maior inteligência
DeepSeek V3.1
2 modelos
14
Maior inteligência
DeepSeek R1 0528 Qwen3 8B
1 modelo
8
Maior inteligência
DeepSeek R1 0528 (May '25)
1 modelo
13
Maior inteligência
DeepSeek V3 0324
1 modelo
10
Maior inteligência
DeepSeek R1 (Jan '25)
1 modelo
11
Maior inteligência
DeepSeek R1 Distill Llama 70B
1 modelo
8
Maior inteligência
DeepSeek R1 Distill Llama 8B
1 modelo
6
Maior inteligência
DeepSeek R1 Distill Qwen 1.5B
1 modelo
6
Maior inteligência
DeepSeek R1 Distill Qwen 14B
1 modelo
8
Maior inteligência
DeepSeek R1 Distill Qwen 32B
1 modelo
8
Maior inteligência
DeepSeek V3 (Dec '24)
1 modelo
8
Maior inteligência
DeepSeek-V2.5 (Dec '24)
1 modelo
7
Maior inteligência
DeepSeek-V2.5
1 modelo
7
Maior inteligência
DeepSeek-Coder-V2
1 modelo
6
Maior inteligência
DeepSeek Coder V2 Lite Instruct
1 modelo
5
Maior inteligência
DeepSeek-V2-Chat
1 modelo
6
Maior inteligência
DeepSeek LLM 67B Chat (V1)
1 modelo
5
Maior inteligência
Mais detalhes
Pesos | Benchmarks de provedores | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flash (Reasoning, Max Effort) | 39 | 552B 16B ativos durante a inferência | 1M | US$ 0,2 | 221 | +15 | |||
| DeepSeek V4 Pro 0813 (Reasoning, Max Effort) | 36 | 1.6T 49B ativos durante a inferência | 1M | US$ 0,7 | 104 | +8 | |||
| DeepSeek V4 Flash Vision (Reasoning, Max Effort) | 35 | 284B 13B ativos durante a inferência | 1M | US$ 0,2 | 223 | Não disponível | +2 | ||
| DeepSeek V4 Flash 0731 (Reasoning, Max Effort) | 34 | 284B 13B ativos durante a inferência | 1M | US$ 0,2 | 227 | +14 | |||
| DeepSeek V4 Pro 0424 (Reasoning, Max Effort) | 30 | 1.6T 49B ativos durante a inferência | 1M | US$ 0,2 | 111 | +7 | |||
| DeepSeek V4 Pro 0424 (Reasoning, High Effort) | 30 | 1.6T 49B ativos durante a inferência | 1M | US$ 0,2 | 106 | +5 | |||
| DeepSeek V4 Flash 0420 (Reasoning, High Effort) | 26 | 284B 13B ativos durante a inferência | 1M | US$ 0,1 | - | +4 | |||
| DeepSeek V4.1 Flash (Non-Reasoning) | 25 | 552B 16B ativos durante a inferência | 1M | US$ 0,2 | 239 | ||||
| DeepSeek V4 Flash 0420 (Reasoning, Max Effort) | 24 | 284B 13B ativos durante a inferência | 1M | US$ 0,1 | - | +2 | |||
| DeepSeek V3.2 (Reasoning) | 21 | 685B 37B ativos durante a inferência | 128k | US$ 0,1 | - | +5 | |||
| DeepSeek V4 Pro 0424 (Non-reasoning) | 21 | 1.6T 49B ativos durante a inferência | 1M | US$ 0,2 | 99 | +3 | |||
| DeepSeek V4 Flash 0420 (Non-reasoning) | 19 | 284B 13B ativos durante a inferência | 1M | US$ 0,1 | - | ||||
| DeepSeek V3.2 Exp (Reasoning) | 17 | 685B 37B ativos durante a inferência | 128k | US$ 0,1 | - | ||||
| DeepSeek V3.2 (Non-reasoning) | 16 | 685B 37B ativos durante a inferência | 128k | US$ 0,3 | - | +8 | |||
| DeepSeek V3.1 Terminus (Reasoning) | 15 | 685B 37B ativos durante a inferência | 128k | US$ 1,7 | - | ||||
| DeepSeek V3.2 Speciale | 14 | 685B 37B ativos durante a inferência | 128k | - | - | - | |||
| DeepSeek V3.1 Terminus (Non-reasoning) | 14 | 685B 37B ativos durante a inferência | 128k | US$ 0,2 | - | ||||
| DeepSeek V3.2 Exp (Non-reasoning) | 14 | 685B 37B ativos durante a inferência | 128k | US$ 0,1 | - | ||||
| DeepSeek V3.1 (Non-reasoning) | 14 | 685B 37B ativos durante a inferência | 128k | US$ 0,7 | - | +5 | |||
| DeepSeek V3.1 (Reasoning) | 13 | 685B 37B ativos durante a inferência | 128k | US$ 0,7 | - | ||||
| DeepSeek R1 0528 (May '25) | 13 | 685B 37B ativos durante a inferência | 128k | US$ 1,5 | - | +2 | |||
| DeepSeek R1 (Jan '25) | 11 | 685B 37B ativos durante a inferência | 128k | US$ 2,2 | - | ||||
| DeepSeek V3 0324 | 10 | 671B 37B ativos durante a inferência | 128k | US$ 0,8 | - | ||||
| DeepSeek V3 (Dec '24) | 8 | 671B 37B ativos durante a inferência | 128k | US$ 0,4 | - | ||||
| DeepSeek R1 Distill Qwen 32B | 8 | 32B | 128k | - | - | - | |||
| DeepSeek R1 0528 Qwen3 8B | 8 | 8.2B | 33k | - | - | - | |||
| DeepSeek R1 Distill Llama 70B | 8 | 70B | 128k | US$ 0,7 | - | ||||
| DeepSeek R1 Distill Qwen 14B | 8 | 14B | 128k | - | - | - | |||
| DeepSeek-V2.5 (Dec '24) | 7 | 236B 21B ativos durante a inferência | 128k | - | - | - | |||
| DeepSeek-V2.5 | 7 | 236B 21B ativos durante a inferência | 128k | - | - | - | |||
| DeepSeek R1 Distill Llama 8B | 6 | 8B | 128k | - | - | - | |||
| DeepSeek-Coder-V2 | 6 | 236B 21B ativos durante a inferência | 128k | - | - | - | |||
| DeepSeek R1 Distill Qwen 1.5B | 6 | 1.5B | 128k | - | - | - | |||
| DeepSeek-V2-Chat | 6 | 236B 21B ativos durante a inferência | 128k | - | - | - | |||
| DeepSeek Coder V2 Lite Instruct | 5 | 16B 2,4B ativos durante a inferência | 128k | - | - | - | |||
| DeepSeek LLM 67B Chat (V1) | 5 | 7B | 4k | - | - | - |