DeepSeek: Models Intelligence, Performance & Price

DeepSeek
DeepSeek

This analysis is intended to support you in choosing the best model provided by DeepSeek for your use-case.

Most Intelligent

Updated
#1
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
39
#2
DeepSeek V4 Pro 0813 (max)DeepSeek V4 Pro 0813 (max)
36
#3
DeepSeek V4 Flash Vision (max)DeepSeek V4 Flash Vision (max)
35
#4
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
34
#5
DeepSeek V4 Pro (max)DeepSeek V4 Pro (max)
30

Intelligence index

Total 9 models

Fastest

#1
DeepSeek V4 Flash 0731 (max)DeepSeek V4 Flash 0731 (max)
215 t/s
#2
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
213 t/s
#3
DeepSeek V4 Flash Vision (max)DeepSeek V4 Flash Vision (max)
211 t/s
#4
DeepSeek V4.1 Flash (non-reasoning)DeepSeek V4.1 Flash (non-reasoning)
208 t/s
#5
DeepSeek V4 Pro (high)DeepSeek V4 Pro (high)
110 t/s

Output speed

Total 9 models

Lowest Price

#1
DeepSeek V4 Pro (max)DeepSeek V4 Pro (max)
$0.18
#2
DeepSeek V4 Pro (high)DeepSeek V4 Pro (high)
$0.18
#3
DeepSeek V4 Pro (non-reasoning)DeepSeek V4 Pro (non-reasoning)
$0.18
#4
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
$0.18
#5
DeepSeek V4.1 Flash (non-reasoning)DeepSeek V4.1 Flash (non-reasoning)
$0.18

Blended price (per 1M tokens)

Total 9 models

DeepSeek offers 9 models, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across models.

  • For intelligence, the top models on DeepSeek are DeepSeek V4.1 Flash (max) (39), DeepSeek V4 Pro 0813 (max) (36), and DeepSeek V4 Flash Vision (max) (35).
  • For output speed, the fastest models are DeepSeek V4 Flash 0731 (max) (215 t/s), DeepSeek V4.1 Flash (max) (213 t/s), and DeepSeek V4 Flash Vision (max) (211 t/s). Speed varies significantly across models, with a 96% difference between the fastest and slowest.
  • For latency, DeepSeek V4.1 Flash (non-reasoning) (0.95s), DeepSeek V4 Pro (non-reasoning) (1.59s), and DeepSeek V4 Pro 0813 (non-reasoning) (1.65s) offer the lowest time to first answer token.
  • For pricing, DeepSeek V4 Pro (max) ($0.18), DeepSeek V4 Pro (high) ($0.18), and DeepSeek V4 Pro (non-reasoning) ($0.18) offer the lowest blended prices per 1M tokens.
  • For context window size, DeepSeek V4 Pro 0813 (max) (1M), DeepSeek V4.1 Flash (max) (1M), and DeepSeek V4 Flash Vision (max) (1M) support the largest context windows on DeepSeek.
Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Intelligence Evaluations

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

SciCodeUnder review

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Pricing

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Performance Summary

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Speed

Measured by Output Speed (tokens per second)

Output Speed

Output tokens per second · Higher is better

Latency

Measured by Time (seconds) to First Token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

End-to-End Response Time

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Further Analysis
DeepSeek logo
DeepSeek V4.1 Flash (max)
1M
Open
39
$0.27
209
0.95
12.89
9.55
DeepSeek logo
DeepSeek V4 Pro 0813 (max)
1.05M
Open
36
$0.67
81
1.67
32.52
24.68
DeepSeek logo
DeepSeek V4 Flash Vision (max)
1M
Proprietary
35
$0.31
215
0.94
12.57
9.31
DeepSeek logo
DeepSeek V4 Flash 0731 (max)
1M
Open
34
$0.22
216
0.97
12.53
9.24
DeepSeek logo
DeepSeek V4 Pro (max)
1M
Open
30
$0.12
84
1.57
59.45
51.94
DeepSeek logo
DeepSeek V4 Pro (high)
1M
Open
30*
--
85
1.69
30.90
23.35
DeepSeek logo
DeepSeek V4.1 Flash (non-reasoning)
1M
Open
25
$0.15
202
1.01
3.48
--
DeepSeek logo
DeepSeek V3.2
128k
Open
21*
--
--
--
--
--
DeepSeek logo
DeepSeek V4 Pro (non-reasoning)
1M
Open
21*
--
78
1.67
8.04
--
DeepSeek logo
DeepSeek V4 Pro 0813 (non-reasoning)
1M
Open
20
$0.36
84
1.95
7.88
--
DeepSeek logo
DeepSeek V3.2 Exp
128k
Open
17*
--
--
--
--
--
DeepSeek logo
DeepSeek V3.2 (non-reasoning)
128k
Open
16*
--
--
--
--
--
DeepSeek logo
DeepSeek V3.2 Exp (non-reasoning)
128k
Open
14*
--
--
--
--
--

Key definitions

Frequently Asked Questions

Common questions about DeepSeek