所有发布

GPT-6 Astra vs. Mistral Small 4:发布比较

GPT-6 Astra 与 Mistral Small 4 发布比较:分别包含 5 个和 2 个模型,各自具有不同的智能、性能与价格特征。下方对全部 7 个模型的关键指标进行了比较。

  • 智能方面,GPT-6 Astra 得分最高:GPT-6 Astra (max)(53),Mistral Small 4 为 Mistral Small 4 (Reasoning)(11)。
  • 输出速度方面,Mistral Small 4 最快:Mistral Small 4 (Reasoning)(181 token/秒),GPT-6 Astra 为 GPT-6 Astra (max)(51 token/秒)。
  • 单任务成本方面,Mistral Small 4 最低:Mistral Small 4 (Reasoning)($0.02),GPT-6 Astra 为 GPT-6 Astra (low)($0.82)。
模型 Playground

比较的发布

智能

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
GPT-6 Astra
Pareto line

Cost per Task (USD, Log Scale)

成本

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

速度与延迟

Output Speed

Output tokens per second · Higher is better

能力得分

能力指数

衡量模型在特定能力与行业中的表现
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better

更多详情

输入模态
输出模态
权重
服务商基准测试
GPT-6 Astra (max)
OpenAI 标志OpenAI
53
55
57
59
52
55
60
$3.26
US$7.7
$10.00
$50.00
$1.00
US$5,324
27k
17k
60M
51
315.42 秒
315.42 秒
325.19 秒
526.58 秒
-
1M
2026年9月
2026年4月
是

支持:文本和图像

支持:文本

不可用
-
Microsoft AzureOpenAIAmazon Bedrock
GPT-6 Astra (xhigh)
OpenAI 标志OpenAI
52
54
57
58
-
55
59
$2.31
US$7.7
$10.00
$50.00
$1.00
US$3,803
17k
9k
38M
48
201.83 秒
201.83 秒
212.17 秒
347.15 秒
-
1M
2026年9月
2026年4月
是

支持:文本和图像

支持:文本

不可用
-
OpenAI
GPT-6 Astra (high)
OpenAI 标志OpenAI
51
53
55
58
-
53
58
$1.73
US$7.7
$10.00
$50.00
$1.00
US$2,925
12k
5k
26M
47
58.39 秒
58.39 秒
69.08 秒
250.59 秒
-
1M
2026年9月
2026年4月
是

支持:文本和图像

支持:文本

不可用
-
OpenAI
GPT-6 Astra (medium)
OpenAI 标志OpenAI
50
52
54
57
-
52
58
$1.54
US$7.7
$10.00
$50.00
$1.00
US$2,434
10k
3k
19M
45
6.21 秒
6.21 秒
17.34 秒
211.08 秒
-
1M
2026年9月
2026年4月
是

支持:文本和图像

支持:文本

不可用
-
OpenAI
GPT-6 Astra (low)
OpenAI 标志OpenAI
46
49
50
53
-
48
55
$0.82
US$7.7
$10.00
$50.00
$1.00
US$1,537
4k
938
10M
44
2.97 秒
2.97 秒
14.28 秒
99.66 秒
-
1M
2026年9月
2026年4月
是

支持:文本和图像

支持:文本

不可用
-
Microsoft AzureOpenAI
Mistral Small 4 (Reasoning)
Mistral 标志Mistral
11
10
7
13
-
16
18
$0.02
US$0.1
$0.15
$0.60
$0.015
US$42
14k
9k
54M
181
0.77 秒
11.84 秒
14.61 秒
91.29 秒
119B
推理时启用 6.5B 个参数
256k
2026年3月
-
是

支持:文本和图像

支持:文本

Mistral
Mistral Small 4 (Non-reasoning)
Mistral 标志Mistral
9
-
-
-
-
-
-
-
US$0.1
$0.15
$0.60
$0.015
-
-
-
-
165
0.77 秒
0.77 秒
3.79 秒
-
119B
推理时启用 6.5B 个参数
256k
2026年3月
-
否

支持:文本和图像

支持:文本

Mistral