此模型已弃用。我们仅继续对默认的 10k 输入 token 工作负载进行性能测试。其他工作负载的结果为历史数据,不再更新。
Anthropic 已推出更新版本 Claude Sonnet 5。我们建议考虑使用该版本。
Claude Sonnet 4.6 (Max) 智能、性能与价格分析
Claude Sonnet 4.6 (Max) 的智能水平高于平均水平,但价格有些昂贵;比较对象为其他价格相近的模型。 此外,其速度低于平均水平,回答也非常冗长。 该模型支持文本和图像输入,可输出文本,上下文窗口为 1M 个 token。
Claude Sonnet 4.6 (Max) 在 Artificial Analysis Intelligence Index 上的得分为 30,在同类模型中高于平均水平(中位数:26)。在 Intelligence Index 评测中,它生成了 230M 个 token;与 81M 的中位数相比,其回答非常冗长。
Claude Sonnet 4.6 (Max) 每 100 万输入 token 的价格为 $3.00(有些昂贵,中位数:$2.00),每 100 万输出 token 的价格为 $15.00(有些昂贵,中位数:$10.00)。在 Intelligence Index 上评测 Claude Sonnet 4.6 (Max) 的平均每项任务成本为 $2.49。
Claude Sonnet 4.6 (Max) 的速度为每秒 59 个 token,低于平均水平(86)。
| 推理 | 是 此页面展示该模型的推理版本。 可能还存在非推理版本。 |
|---|---|
| 输入模态 | 支持:文本和图像 |
| 输出模态 | 支持:文本 |
| 上下文窗口 | 1M 约 1500 页 A4 纸(12 号 Arial 字体) |
指标与同类别模型进行比较:
- 非推理模型 → 仅与其他非推理模型比较
- 推理模型 → 同时与推理和非推理模型比较
- 开放权重模型 → 仅与规模类别相同的其他开放权重模型比较:
- 微型:≤4B 个参数
- 小型:4B–40B 个参数
- 中型:40B–150B 个参数
- 大型:>150B 个参数
- 专有模型 → 与价格区间相同的专有模型和开放权重模型比较,采用输入/输出价格 3:1 的混合比例:
- 每 100 万 token <$0.15
- 每 100 万 token $0.15–$1
- 每 100 万 token >$1
智能Updated
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index:开放权重与专有模型
衡量模型在特定能力和行业中的表现
Artificial Analysis 财务与会计指数
Intelligence Evaluations
Agentic knowledge work, (Elo-500)/2000
Agentic real-world work tasks, (Elo-500)/2000
Agentic SaaS workflows
Agentic coding & terminal use
Coding
Reasoning & knowledge
Professional document reasoning, All-pass
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Legal agentic work, criterion pass rate
Agentic business operations
Agentic scientific research workflows in a terminal
Quantitative analysis on spreadsheets & documents
Kubernetes incident root-cause analysis
Visual reasoning
Medical long context reasoning
AA-Briefcase v1.1Updated
AA-Briefcase Elo
AA-Omniscience
AA-Omniscience Index
Intelligence Index 比较
Intelligence Index 与每项任务的成本
Token 使用量
智能指数每项任务的输出 Token 数
成本
每项 Intelligence Index 任务的成本
运行 Artificial Analysis Intelligence Index 的成本
价格:缓存命中、输入和输出
上下文窗口
上下文窗口
速度
按输出速度(每秒 token 数)衡量
输出速度
智能指数每项任务耗时
延迟
按首 Token 延迟(秒)衡量
延迟: 首个答案 Token 延迟
端到端响应时间
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
端到端响应时间
常见问题
关于 Claude Sonnet 4.6 (Max) 的常见问题
Claude Sonnet 4.6 (Max) 发布于 2026年2月17日。
Claude Sonnet 4.6 (Max) 由 Anthropic 开发。
Claude Sonnet 4.6 (Max) 在 Artificial Analysis Intelligence Index 上的得分为 30,在其他价格档位相近的推理模型中高于平均水平(中位数:26)。
Claude Sonnet 4.6 (Max) 以每秒 59.1 个 token 的速度生成输出(基于 Anthropic 的 API),与其他价格档位相近的推理模型相比低于平均水平(中位数:86.1 t/s)。
Claude Sonnet 4.6 (Max) 的首 Token 延迟(TTFT)为 126.44 秒(基于 Anthropic 的 API),与其他价格档位相近的推理模型相比处于较高水平(中位数:3.83 秒)。
Claude Sonnet 4.6 (Max) 每 100 万输入 token 的价格为 $3.00(略高于平均水平,中位数:$2.00),每 100 万输出 token 的价格为 $15.00(略高于平均水平,中位数:$10.00);基于 Anthropic 的 API。
Claude Sonnet 4.6 (Max) 每 100 万输入 token 的价格为 $3.00,每 100 万输出 token 的价格为 $15.00(基于 Anthropic 的 API)。按缓存命中/输入/输出为 7:2:1 的比例计算,混合价格为每 100 万 token $2.31。价格可能因服务商而异。 比较服务商价格
在 Intelligence Index 评测中,Claude Sonnet 4.6 (Max) 生成了 230M 个输出 token;与其他价格档位相近的推理模型相比处于较高水平(中位数:81M)。
是,Claude Sonnet 4.6 (Max) 是推理模型。它会在给出答案之前,通过扩展思考或思维链推理来解决复杂问题。
Claude Sonnet 4.6 (Max) 支持文本和图像输入。
Claude Sonnet 4.6 (Max) 支持文本输出。
是,Claude Sonnet 4.6 (Max) 支持图像输入,可以分析、描述图像并回答有关图像的问题。
是,Claude Sonnet 4.6 (Max) 是多模态模型,可以处理文本和图像输入并生成文本输出。
Claude Sonnet 4.6 (Max) 的上下文窗口为 1.0M 个 token。这决定了模型在单次请求中可以处理多少文本和对话历史。
否,Claude Sonnet 4.6 (Max) 是专有模型,其模型权重并未公开。
Claude Sonnet 4.6 (Max) 是专有模型,Anthropic 尚未披露模型规模或参数量。
Claude Sonnet 4.6 (Max) 在 Artificial Analysis Intelligence Index 上的得分为 30。这项综合基准测试评估模型的推理、知识、数学和编程能力。
是,可通过 4 家服务商的 API 使用 Claude Sonnet 4.6 (Max)。 比较 API 服务商
可通过 4 家 API 服务商使用 Claude Sonnet 4.6 (Max)。 比较服务商