This model is deprecated. We only continue performance benchmarking for the default 10k input token workload. Results for other workloads are historical and no longer updated.
Alibaba has launched a newer model, Qwen3.5 0.8B. We suggest considering it instead.
For more information, see comparison of Qwen3.5 0.8B to other models and API provider benchmarks for Qwen3.5 0.8B.
Qwen3 0.6B (Non-reasoning) APIプロバイダーのベンチマークと分析
この分析は、ユースケースに最適なQwen3 0.6B (Non-reasoning)のAPIプロバイダーを選ぶための参考情報です。
最速
出力速度
プロバイダー合計:0社
最短の遅延
最初のトークンまでの時間
プロバイダー合計:0社
最安料金
ブレンド料金(100万トークンあたり)
プロバイダー合計:0社
現在、Qwen3 0.6Bを利用できるAPIプロバイダーはありません。
モデルの詳細と他のモデルとの知能比較は、Qwen3 0.6B (Non-reasoning)のモデルページをご覧ください。
ハイライト
料金
Pricing: Cache Hit, Input, and Output
Pricing: Blended Price
Output Speed vs. Price
速度
出力速度(1秒あたりのトークン数)で測定
Output Speed: Qwen3 0.6B (Non-reasoning)
Latency vs. Output Speed
遅延
最初のトークンまでの時間(秒)で測定
最初のトークンまでの時間
エンドツーエンド応答時間
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
エンドツーエンド応答時間
主要な比較指標とAPI機能
| No results. | |||||||||
よくある質問
Qwen3 0.6B (Non-reasoning)のプロバイダーに関するよくある質問
現在、ベンチマーク対象のAPIプロバイダーでQwen3 0.6B (Non-reasoning)を利用することはできません。オープンウェイトモデルのため、セルフホストできます。