发布比较工具

并排比较最多五个模型发布,覆盖每个发布的所有变体,包括智能、价格、输出速度、延迟、上下文窗口等指标。

如需了解我们如何衡量和评估模型,请参阅方法论页面

比较的发布

智能

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Claude Fable 5.1
GPT-6 Astra
Pareto line

成本

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

速度与延迟

Output Speed

Output tokens per second · Higher is better

能力得分

能力指数

衡量模型在特定能力与行业中的表现
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench v4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better

更多详情

输入模态
输出模态
权重
服务商基准测试
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
Anthropic 标志Anthropic
53
57
60
61
58
57
63
$7.63
US$7.2
$10.00
$50.00
$0.25
US$13,129
78k
47k
188M
67
212.01 秒
212.01 秒
219.46 秒
737.06 秒
-
1M
2026年9月
-

支持:文本和图像

支持:文本

不可用
-
Anthropic
Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)
Anthropic 标志Anthropic
53
56
58
61
-
57
62
$5.98
US$7.2
$10.00
$50.00
$0.25
US$9,063
61k
34k
121M
62
85.94 秒
85.94 秒
93.96 秒
628.62 秒
-
1M
2026年9月
-

支持:文本和图像

支持:文本

不可用
-
Anthropic
GPT-6 Astra (max)
OpenAI 标志OpenAI
53
55
58
59
52
55
60
$3.26
US$7.7
$10.00
$50.00
$1.00
US$5,324
27k
17k
60M
60
294.93 秒
294.93 秒
303.25 秒
451.36 秒
-
1M
2026年9月
2026年4月

支持:文本和图像

支持:文本

不可用
-
OpenAI
GPT-6 Astra (xhigh)
OpenAI 标志OpenAI
53
55
57
58
-
55
60
$2.31
US$7.7
$10.00
$50.00
$1.00
US$3,803
17k
9k
38M
55
125.35 秒
125.35 秒
134.45 秒
308.44 秒
-
1M
2026年9月
2026年4月

支持:文本和图像

支持:文本

不可用
-
OpenAI
Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)
Anthropic 标志Anthropic
51
54
56
58
-
55
60
$3.91
US$7.2
$10.00
$50.00
$0.25
US$5,242
38k
19k
62M
57
17.77 秒
17.77 秒
26.49 秒
420.30 秒
-
1M
2026年9月
-

支持:文本和图像

支持:文本

不可用
-
Anthropic
GPT-6 Astra (high)
OpenAI 标志OpenAI
51
53
56
58
-
53
59
$1.72
US$7.7
$10.00
$50.00
$1.00
US$2,917
12k
5k
26M
52
36.82 秒
36.82 秒
46.35 秒
227.77 秒
-
1M
2026年9月
2026年4月

支持:文本和图像

支持:文本

不可用
-
OpenAI
GPT-6 Astra (medium)
OpenAI 标志OpenAI
50
52
54
57
-
52
58
$1.54
US$7.7
$10.00
$50.00
$1.00
US$2,434
10k
3k
19M
51
3.96 秒
3.96 秒
13.75 秒
188.62 秒
-
1M
2026年9月
2026年4月

支持:文本和图像

支持:文本

不可用
-
OpenAI
Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback)
Anthropic 标志Anthropic
49
52
54
57
-
52
58
$2.98
US$7.2
$10.00
$50.00
$0.25
US$3,983
28k
12k
44M
56
7.88 秒
7.88 秒
16.76 秒
311.81 秒
-
1M
2026年9月
-

支持:文本和图像

支持:文本

不可用
-
Anthropic
Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback)
Anthropic 标志Anthropic
47
49
52
54
-
49
55
$2.37
US$7.2
$10.00
$50.00
$0.25
US$3,158
22k
8k
33M
57
4.73 秒
4.73 秒
13.58 秒
238.57 秒
-
1M
2026年9月
-

支持:文本和图像

支持:文本

不可用
-
Anthropic
GPT-6 Astra (low)
OpenAI 标志OpenAI
46
49
50
53
-
48
55
$0.82
US$7.7
$10.00
$50.00
$1.00
US$1,537
4k
938
10M
52
2.40 秒
2.40 秒
11.94 秒
84.35 秒
-
1M
2026年9月
2026年4月

支持:文本和图像

支持:文本

不可用
-
OpenAI