该模型暂无服务商基准测试数据。
有关模型详情及其与其他模型的智能比较,请参阅 Granite 4.0 H 1B 的模型页面。
亮点
更新:默认性能基准测试工作负载已调整为 1 万个输入 token,以更贴近生产使用场景。你仍可在上方选择其他工作负载。
价格
Pricing: Cache Hit, Input, and Output
Price (USD per M Tokens) · Lower is better · 10,000 input tokens
No data available
Pricing: Blended Price
Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended) · Lower is better
No data available
Output Speed vs. Price
Blended at 7:2:1 (cache-input-output) · Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
No data available
速度
按输出速度(每秒 token 数)衡量
Output Speed: Granite 4.0 H 1B
Output speed: output tokens per second · 10,000 input tokens
No data available
Latency vs. Output Speed
Latency: seconds to first token received · Output speed: output tokens per second · 10,000 input tokens
Most attractive quadrant
No data available
延迟
按首 Token 延迟(秒)衡量
首 Token 延迟:Granite 4.0 H 1B 服务商
Seconds to first token received · Lower is better · 10,000 input tokens
No data available
端到端响应时间
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
端到端响应时间:Granite 4.0 H 1B 服务商
Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better · 10,000 input tokens
No data available
关键比较指标与 API 功能
| No results. | |||||||||
常见问题
关于 Granite 4.0 H 1B 服务商的常见问题
目前我们进行基准测试的 API 服务商均未提供 Granite 4.0 H 1B。该模型为开放权重模型,可以自行托管。