This model is deprecated. We only continue performance benchmarking for the default 10k input token workload. Results for other workloads are historical and no longer updated.
Alibaba has launched a newer model, Qwen3 4B 2507. We suggest considering it instead.
For more information, see comparison of Qwen3 4B 2507 to other models and API provider benchmarks for Qwen3 4B 2507.
本分析旨在帮助你根据使用场景,选择 Qwen3 4B (Non-reasoning) 的最佳 API 服务商。
速度最快
输出速度
共 0 家服务商
延迟最低
首 Token 延迟
共 0 家服务商
价格最低
每 100 万 token 的混合价格
共 0 家服务商
目前没有可用的 Qwen3 4B API 服务商。
有关模型详情及其与其他模型的智能比较,请参阅 Qwen3 4B (Non-reasoning) 的模型页面。
亮点
价格
Pricing: Cache Hit, Input, and Output
Pricing: Blended Price
Output Speed vs. Price
速度
按输出速度(每秒 token 数)衡量
Output Speed: Qwen3 4B (Non-reasoning)
Latency vs. Output Speed
延迟
按首 Token 延迟(秒)衡量
首 Token 延迟
端到端响应时间
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
端到端响应时间
关键比较指标与 API 功能
| No results. | |||||||||
常见问题
关于 Qwen3 4B (Non-reasoning) 服务商的常见问题
目前我们进行基准测试的 API 服务商均未提供 Qwen3 4B (Non-reasoning)。该模型为开放权重模型,可以自行托管。