This model is deprecated. We only continue performance benchmarking for the default 10k input token workload. Results for other workloads are historical and no longer updated.
NVIDIA has launched a newer model, Llama Nemotron Super 49B v1.5. We suggest considering it instead.
For more information, see comparison of Llama Nemotron Super 49B v1.5 to other models and API provider benchmarks for Llama Nemotron Super 49B v1.5.
Llama 3.3 Nemotron Super 49B v1 (Reasoning) API 服务商基准测试与分析
本分析旨在帮助你根据使用场景,选择 Llama 3.3 Nemotron Super 49B v1 (Reasoning) 的最佳 API 服务商。
速度最快
输出速度
共 0 家服务商
延迟最低
首个答案 Token 延迟
共 0 家服务商
价格最低
每 100 万 token 的混合价格
共 0 家服务商
目前没有可用的 Llama 3.3 Nemotron Super 49B API 服务商。
有关模型详情及其与其他模型的智能比较,请参阅 Llama 3.3 Nemotron Super 49B v1 (Reasoning) 的模型页面。
亮点
价格
Pricing: Cache Hit, Input, and Output
Pricing: Blended Price
Output Speed vs. Price
速度
按输出速度(每秒 token 数)衡量
Output Speed: Llama 3.3 Nemotron Super 49B v1 (Reasoning)
Latency vs. Output Speed
延迟
按首 Token 延迟(秒)衡量
首个答案 Token 延迟
端到端响应时间
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
端到端响应时间
关键比较指标与 API 功能
| No results. | |||||||||
常见问题
关于 Llama 3.3 Nemotron Super 49B v1 (Reasoning) 服务商的常见问题
目前我们进行基准测试的 API 服务商均未提供 Llama 3.3 Nemotron Super 49B v1 (Reasoning)。该模型为开放权重模型,可以自行托管。