| Claude Opus 5.5 (max with fallback) | Claude Opus 5.5 (xhigh with fallback) | Claude Opus 5.5 (high with fallback) | GPT-6 Astra (max) | GPT-6 Astra (xhigh) | Claude Opus 5.5 (medium with fallback) | GPT-6 Astra (high) | GPT-6 Astra (medium) | GPT-6 Astra (low) | Claude Opus 5.5 (low with fallback) | |
|---|---|---|---|---|---|---|---|---|---|---|
| Intelligence Index | 58 | 56 | 54 | 53 | 52 | 51 | 51 | 50 | 46 | 42 |
| 财务与会计 | 61 | 58 | 56 | 55 | 54 | 54 | 53 | 52 | 49 | 46 |
| 战略与运营 | 64 | 62 | 59 | 57 | 57 | 57 | 55 | 54 | 50 | 48 |
| 法律 | 63 | 61 | 59 | 59 | 58 | 57 | 58 | 57 | 53 | 52 |
| 医疗与健康 | 61 | -- | -- | 52 | -- | -- | -- | -- | -- | -- |
| 每任务成本美元 | $5.98 | $3.46 | $1.82 | $3.26 | $2.31 | $1.34 | $1.73 | $1.54 | $0.82 | $0.55 |
| 输出速度Token/秒 | 92 | 76 | 72 | 57 | 54 | 72 | 51 | 49 | 49 | 70 |
| 首 Token 延迟秒 | 723.15 | 147.22 | 34.40 | 300.30 | 142.98 | 26.57 | 68.15 | 4.52 | 2.65 | 6.31 |
| 上下文窗口 | 1M | 1M | 1M | 1M | 1M | 1M | 1M | 1M | 1M | 1M |
| 开发商 | ||||||||||
| 许可证 | 专有 | 专有 | 专有 | 专有 | 专有 | 专有 | 专有 | 专有 | 专有 | 专有 |
| 输入模态 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 | 支持:文本和图像 |
| 输出模态 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 | 支持:文本 |
| 服务商 | ||||||||||
智能
Artificial Analysis Intelligence Index
Intelligence Index 与每项任务的成本
单任务成本(美元,对数刻度)
成本
每项 Intelligence Index 任务的成本
速度与延迟
输出速度
能力得分
能力指数
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better