此模型已弃用。我们仅继续对默认的 10k 输入 token 工作负载进行性能测试。其他工作负载的结果为历史数据,不再更新。
Anthropic 已推出更新版本 Claude Sonnet 5.5。我们建议考虑使用该版本。
Claude Sonnet 5 (Xhigh) 智能、性能与价格分析
Claude Sonnet 5 (Xhigh) 的智能水平高于平均水平,价格也较为合理;比较对象为其他价格相近的模型。 此外,其速度低于平均水平,回答也非常冗长。 该模型支持文本和图像输入,可输出文本,上下文窗口为 1M 个 token。
Claude Sonnet 5 (Xhigh) 在 Artificial Analysis Intelligence Index 上的得分为 34,在同类模型中高于平均水平(中位数:26)。在 Intelligence Index 评测中,它生成了 130M 个 token;与 81M 的中位数相比,其回答非常冗长。
Claude Sonnet 5 (Xhigh) 每 100 万输入 token 的价格为 $2.00(中等,中位数:$2.00),每 100 万输出 token 的价格为 $10.00(中等,中位数:$10.00)。在 Intelligence Index 上评测 Claude Sonnet 5 (Xhigh) 的平均每项任务成本为 $2.87。
Claude Sonnet 5 (Xhigh) 的速度为每秒 72 个 token,低于平均水平(79)。
| 推理 | 是 此页面展示该模型的推理版本。 可能还存在非推理版本。 |
|---|---|
| 输入模态 | 支持:文本和图像 |
| 输出模态 | 支持:文本 |
| 上下文窗口 | 1M 约 1500 页 A4 纸(12 号 Arial 字体) |
指标与同类别模型进行比较:
- 非推理模型 → 仅与其他非推理模型比较
- 推理模型 → 同时与推理和非推理模型比较
- 开放权重模型 → 仅与规模类别相同的其他开放权重模型比较:
- 微型:≤4B 个参数
- 小型:4B–40B 个参数
- 中型:40B–150B 个参数
- 大型:>150B 个参数
- 专有模型 → 与价格区间相同的专有模型和开放权重模型比较,采用输入/输出价格 3:1 的混合比例:
- 每 100 万 token <$0.15
- 每 100 万 token $0.15–$1
- 每 100 万 token >$1
智能Updated
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index:开放权重与专有模型
衡量模型在特定能力和行业中的表现
Artificial Analysis 财务与会计指数
Intelligence Evaluations
Agentic knowledge work, (Elo-500)/2000
Agentic real-world work tasks, (Elo-500)/2000
Agentic SaaS workflows
Agentic coding & terminal use
Coding
Reasoning & knowledge
Professional document reasoning, All-pass
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Legal agentic work, criterion pass rate
Agentic business operations
Agentic scientific research workflows in a terminal
Quantitative analysis on spreadsheets & documents
Kubernetes incident root-cause analysis
Visual reasoning
Medical long context reasoning
AA-Briefcase v1.1Updated
AA-Briefcase Elo
AA-Omniscience
AA-Omniscience Index
Intelligence Index 比较
Intelligence Index 与每项任务的成本
Token 使用量
智能指数每项任务的输出 Token 数
成本
每项 Intelligence Index 任务的成本
运行 Artificial Analysis Intelligence Index 的成本
价格:缓存命中、输入和输出
上下文窗口
上下文窗口
速度
按输出速度(每秒 token 数)衡量
输出速度
智能指数每项任务耗时
延迟
按首 Token 延迟(秒)衡量
延迟: 首个答案 Token 延迟
端到端响应时间
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
端到端响应时间
常见问题
关于 Claude Sonnet 5 (Xhigh) 的常见问题
Claude Sonnet 5 (Xhigh) 发布于 2026年6月30日。
Claude Sonnet 5 (Xhigh) 由 Anthropic 开发。
Claude Sonnet 5 (Xhigh) 在 Artificial Analysis Intelligence Index 上的得分为 34,在其他价格档位相近的推理模型中高于平均水平(中位数:26)。
Claude Sonnet 5 (Xhigh) 以每秒 72.3 个 token 的速度生成输出(基于 Anthropic 的 API),与其他价格档位相近的推理模型相比低于平均水平(中位数:78.8 t/s)。
Claude Sonnet 5 (Xhigh) 的首 Token 延迟(TTFT)为 22.89 秒(基于 Anthropic 的 API),与其他价格档位相近的推理模型相比处于较高水平(中位数:3.81 秒)。
Claude Sonnet 5 (Xhigh) 每 100 万输入 token 的价格为 $2.00(优于平均水平,中位数:$2.00),每 100 万输出 token 的价格为 $10.00(优于平均水平,中位数:$10.00);基于 Anthropic 的 API。
Claude Sonnet 5 (Xhigh) 每 100 万输入 token 的价格为 $2.00,每 100 万输出 token 的价格为 $10.00(基于 Anthropic 的 API)。按缓存命中/输入/输出为 7:2:1 的比例计算,混合价格为每 100 万 token $1.54。价格可能因服务商而异。 比较服务商价格
在 Intelligence Index 评测中,Claude Sonnet 5 (Xhigh) 生成了 130M 个输出 token;与其他价格档位相近的推理模型相比略高于平均水平(中位数:81M)。
是,Claude Sonnet 5 (Xhigh) 是推理模型。它会在给出答案之前,通过扩展思考或思维链推理来解决复杂问题。
Claude Sonnet 5 (Xhigh) 支持文本和图像输入。
Claude Sonnet 5 (Xhigh) 支持文本输出。
是,Claude Sonnet 5 (Xhigh) 支持图像输入,可以分析、描述图像并回答有关图像的问题。
是,Claude Sonnet 5 (Xhigh) 是多模态模型,可以处理文本和图像输入并生成文本输出。
Claude Sonnet 5 (Xhigh) 的上下文窗口为 1.0M 个 token。这决定了模型在单次请求中可以处理多少文本和对话历史。
否,Claude Sonnet 5 (Xhigh) 是专有模型,其模型权重并未公开。
Claude Sonnet 5 (Xhigh) 是专有模型,Anthropic 尚未披露模型规模或参数量。
Claude Sonnet 5 (Xhigh) 在 Artificial Analysis Intelligence Index 上的得分为 34。这项综合基准测试评估模型的推理、知识、数学和编程能力。
是,可通过 1 家服务商的 API 使用 Claude Sonnet 5 (Xhigh)。 比较 API 服务商
可通过 1 家 API 服务商使用 Claude Sonnet 5 (Xhigh)。 比较服务商