LithosAI:模型智能、性能与价格

LithosAI
LithosAI

本分析旨在帮助你根据使用场景,选择 LithosAI 提供的最佳模型。

最智能

Updated
#1
Kimi K3 (max) (ULTRA CHAT)Kimi K3 (max) (ULTRA CHAT)
44
#2
Kimi K3 (max)Kimi K3 (max)
44
#3
Kimi K3 (max) (FAST)Kimi K3 (max) (FAST)
44
#4
Kimi K3 (max) (ULTRA)Kimi K3 (max) (ULTRA)
44
#5
DeepSeek V4.1 Flash (max) (FAST)DeepSeek V4.1 Flash (max) (FAST)
39

Intelligence Index

共 11 个模型

速度最快

#1
DeepSeek V4.1 Flash (max) (ULTRA CHAT)DeepSeek V4.1 Flash (max) (ULTRA CHAT)
607 t/s
#2
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)
579 t/s
#3
DeepSeek V4.1 Flash (max) (ULTRA)DeepSeek V4.1 Flash (max) (ULTRA)
566 t/s
#4
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)
562 t/s
#5
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
534 t/s

输出速度

共 11 个模型

价格最低

#1
DeepSeek V4.1 Flash (max)DeepSeek V4.1 Flash (max)
$0.09
#2
DeepSeek V4.1 Flash (max) (FAST)DeepSeek V4.1 Flash (max) (FAST)
$0.15
#3
DeepSeek V4.1 Flash (Non-reasoning) (FAST)DeepSeek V4.1 Flash (Non-reasoning) (FAST)
$0.15
#4
DeepSeek V4.1 Flash (max) (ULTRA)DeepSeek V4.1 Flash (max) (ULTRA)
$0.21
#5
DeepSeek V4.1 Flash (max) (ULTRA CHAT)DeepSeek V4.1 Flash (max) (ULTRA CHAT)
$0.21

每 100 万 token 的混合价格

共 11 个模型

LithosAI 提供 11 个模型,每个模型的智能、性能和价格特征各不相同。 下方对比了各模型的关键指标。

  • 智能方面,LithosAI 上表现最好的模型是 Kimi K3 (max) (ULTRA CHAT)(44)、Kimi K3 (max)(44)和Kimi K3 (max) (FAST)(44)。
  • 输出速度方面,最快的模型是 DeepSeek V4.1 Flash (max) (ULTRA CHAT)(607 t/s)、DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)(579 t/s)和DeepSeek V4.1 Flash (max) (ULTRA)(566 t/s)。
  • 延迟方面,DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)(0.79 秒)、DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)(1.01 秒)和DeepSeek V4.1 Flash (Non-reasoning) (FAST)(1.08 秒) 的首个答案 Token 延迟最低。
  • 价格方面,DeepSeek V4.1 Flash (max)($0.09)、DeepSeek V4.1 Flash (max) (FAST)($0.15)和DeepSeek V4.1 Flash (Non-reasoning) (FAST)($0.15) 每 100 万 token 的混合价格最低。 各模型价格最多相差 2.3 倍。
  • 上下文窗口方面,Kimi K3 (max) (ULTRA CHAT)(1M)、Kimi K3 (max)(1M)和Kimi K3 (max) (FAST)(1M) 支持 LithosAI 上最大的上下文窗口。

亮点

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

智能评测

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Agentic tool use

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

上下文窗口

Context Window

Context window: tokens limit · Higher is better

价格

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

性能摘要

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

速度

按输出速度(每秒 token 数)衡量

Output Speed

Output tokens per second · Higher is better

延迟

按首 Token 延迟(秒)衡量

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

端到端响应时间

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

进一步分析
Kimi 标志
Kimi K3 (max) (ULTRA CHAT)
1.05M
开放
44
$3.93
289
0.71
9.37
6.93
Kimi 标志
Kimi K3 (max)
1.05M
开放
44
$1.69
216
0.71
12.28
9.26
Kimi 标志
Kimi K3 (max) (FAST)
1.05M
开放
44
$2.81
212
0.76
12.57
9.45
Kimi 标志
Kimi K3 (max) (ULTRA)
1.05M
开放
44
$3.93
221
0.75
12.06
9.05
DeepSeek 标志
DeepSeek V4.1 Flash (max) (FAST)
1.05M
开放
39
$0.27
520
1.43
6.23
3.84
DeepSeek 标志
DeepSeek V4.1 Flash (max) (ULTRA)
1.05M
开放
39
$0.38
566
1.13
5.55
3.53
DeepSeek 标志
DeepSeek V4.1 Flash (max) (ULTRA CHAT)
1.05M
开放
39
$0.38
607
0.81
4.93
3.29
DeepSeek 标志
DeepSeek V4.1 Flash (max)
1.05M
开放
39
$0.16
534
1.09
5.78
3.74
DeepSeek 标志
DeepSeek V4.1 Flash (Non-reasoning) (FAST)
1.05M
开放
25
$0.16
511
1.08
2.06
--
DeepSeek 标志
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA)
1.05M
开放
25
$0.22
579
1.01
1.87
--
DeepSeek 标志
DeepSeek V4.1 Flash (Non-reasoning) (ULTRA CHAT)
1.05M
开放
25
$0.22
562
0.79
1.68
--

关键定义

常见问题

关于 LithosAI 的常见问题