DeepSeek Models: Intelligence, Performance & Price
Artificial Analysis Intelligence Index
Artificial Analysis has benchmarked 36 models from DeepSeek. Below is a comparison of the key metrics across these models.
- For intelligence, the top model from DeepSeek is DeepSeek V4.1 Flash (Reasoning, Max Effort) at 39.
- For output speed, the fastest model is DeepSeek V4.1 Flash (Non-Reasoning) at 239 t/s.
- For latency, DeepSeek V4.1 Flash (Non-Reasoning) at 1.12s offers the lowest time to first answer token.
- For pricing, DeepSeek V4.1 Flash (Non-Reasoning) at $0.15 offers the lowest cost per task. Prices vary up to 4.6x across models.
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
All DeepSeek Releases
DeepSeek V4.1 Flash
2 models
39
Highest Intelligence
DeepSeek V4 Flash Vision
1 model
35
Highest Intelligence
DeepSeek V4 Pro 0813
1 model
36
Highest Intelligence
DeepSeek V4 Flash 0731
1 model
34
Highest Intelligence
DeepSeek V4 Flash 0420
3 models
26
Highest Intelligence
DeepSeek V4 Pro 0424
3 models
30
Highest Intelligence
DeepSeek V3.2
2 models
21
Highest Intelligence
DeepSeek V3.2 Speciale
1 model
14
Highest Intelligence
DeepSeek V3.2 Exp
2 models
17
Highest Intelligence
DeepSeek V3.1 Terminus
2 models
15
Highest Intelligence
DeepSeek V3.1
2 models
14
Highest Intelligence
DeepSeek R1 0528 Qwen3 8B
1 model
8
Highest Intelligence
DeepSeek R1 0528 (May '25)
1 model
13
Highest Intelligence
DeepSeek V3 0324
1 model
10
Highest Intelligence
DeepSeek R1 (Jan '25)
1 model
11
Highest Intelligence
DeepSeek R1 Distill Llama 70B
1 model
8
Highest Intelligence
DeepSeek R1 Distill Llama 8B
1 model
6
Highest Intelligence
DeepSeek R1 Distill Qwen 1.5B
1 model
6
Highest Intelligence
DeepSeek R1 Distill Qwen 14B
1 model
8
Highest Intelligence
DeepSeek R1 Distill Qwen 32B
1 model
8
Highest Intelligence
DeepSeek V3 (Dec '24)
1 model
8
Highest Intelligence
DeepSeek-V2.5 (Dec '24)
1 model
7
Highest Intelligence
DeepSeek-V2.5
1 model
7
Highest Intelligence
DeepSeek-Coder-V2
1 model
6
Highest Intelligence
DeepSeek Coder V2 Lite Instruct
1 model
5
Highest Intelligence
DeepSeek-V2-Chat
1 model
6
Highest Intelligence
DeepSeek LLM 67B Chat (V1)
1 model
5
Highest Intelligence
Further details
Weights | Provider Benchmarks | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flash (Reasoning, Max Effort) | 39 | 552B 16B active at inference time | 1M | $0.2 | 221 | +15 | |||
| DeepSeek V4 Pro 0813 (Reasoning, Max Effort) | 36 | 1.6T 49B active at inference time | 1M | $0.7 | 104 | +8 | |||
| DeepSeek V4 Flash Vision (Reasoning, Max Effort) | 35 | 284B 13B active at inference time | 1M | $0.2 | 223 | Not available | +2 | ||
| DeepSeek V4 Flash 0731 (Reasoning, Max Effort) | 34 | 284B 13B active at inference time | 1M | $0.2 | 227 | +14 | |||
| DeepSeek V4 Pro 0424 (Reasoning, Max Effort) | 30 | 1.6T 49B active at inference time | 1M | $0.2 | 111 | +7 | |||
| DeepSeek V4 Pro 0424 (Reasoning, High Effort) | 30 | 1.6T 49B active at inference time | 1M | $0.2 | 106 | +5 | |||
| DeepSeek V4 Flash 0420 (Reasoning, High Effort) | 26 | 284B 13B active at inference time | 1M | $0.1 | - | +4 | |||
| DeepSeek V4.1 Flash (Non-Reasoning) | 25 | 552B 16B active at inference time | 1M | $0.2 | 239 | ||||
| DeepSeek V4 Flash 0420 (Reasoning, Max Effort) | 24 | 284B 13B active at inference time | 1M | $0.1 | - | +2 | |||
| DeepSeek V3.2 (Reasoning) | 21 | 685B 37B active at inference time | 128k | $0.1 | - | +5 | |||
| DeepSeek V4 Pro 0424 (Non-reasoning) | 21 | 1.6T 49B active at inference time | 1M | $0.2 | 99 | +3 | |||
| DeepSeek V4 Flash 0420 (Non-reasoning) | 19 | 284B 13B active at inference time | 1M | $0.1 | - | ||||
| DeepSeek V3.2 Exp (Reasoning) | 17 | 685B 37B active at inference time | 128k | $0.1 | - | ||||
| DeepSeek V3.2 (Non-reasoning) | 16 | 685B 37B active at inference time | 128k | $0.3 | - | +8 | |||
| DeepSeek V3.1 Terminus (Reasoning) | 15 | 685B 37B active at inference time | 128k | $1.7 | - | ||||
| DeepSeek V3.2 Speciale | 14 | 685B 37B active at inference time | 128k | - | - | - | |||
| DeepSeek V3.1 Terminus (Non-reasoning) | 14 | 685B 37B active at inference time | 128k | $0.2 | - | ||||
| DeepSeek V3.2 Exp (Non-reasoning) | 14 | 685B 37B active at inference time | 128k | $0.1 | - | ||||
| DeepSeek V3.1 (Non-reasoning) | 14 | 685B 37B active at inference time | 128k | $0.7 | - | +5 | |||
| DeepSeek V3.1 (Reasoning) | 13 | 685B 37B active at inference time | 128k | $0.7 | - | ||||
| DeepSeek R1 0528 (May '25) | 13 | 685B 37B active at inference time | 128k | $1.5 | - | +2 | |||
| DeepSeek R1 (Jan '25) | 11 | 685B 37B active at inference time | 128k | $2.2 | - | ||||
| DeepSeek V3 0324 | 10 | 671B 37B active at inference time | 128k | $0.8 | - | ||||
| DeepSeek V3 (Dec '24) | 8 | 671B 37B active at inference time | 128k | $0.4 | - | ||||
| DeepSeek R1 Distill Qwen 32B | 8 | 32B | 128k | - | - | - | |||
| DeepSeek R1 0528 Qwen3 8B | 8 | 8.2B | 33k | - | - | - | |||
| DeepSeek R1 Distill Llama 70B | 8 | 70B | 128k | $0.7 | - | ||||
| DeepSeek R1 Distill Qwen 14B | 8 | 14B | 128k | - | - | - | |||
| DeepSeek-V2.5 (Dec '24) | 7 | 236B 21B active at inference time | 128k | - | - | - | |||
| DeepSeek-V2.5 | 7 | 236B 21B active at inference time | 128k | - | - | - | |||
| DeepSeek R1 Distill Llama 8B | 6 | 8B | 128k | - | - | - | |||
| DeepSeek-Coder-V2 | 6 | 236B 21B active at inference time | 128k | - | - | - | |||
| DeepSeek R1 Distill Qwen 1.5B | 6 | 1.5B | 128k | - | - | - | |||
| DeepSeek-V2-Chat | 6 | 236B 21B active at inference time | 128k | - | - | - | |||
| DeepSeek Coder V2 Lite Instruct | 5 | 16B 2.4B active at inference time | 128k | - | - | - | |||
| DeepSeek LLM 67B Chat (V1) | 5 | 7B | 4k | - | - | - |