Llama 3.2 Instruct 90B (Vision) APIプロバイダーのベンチマークと分析
Llama 3.2 Instruct 90B (Vision)のAPIプロバイダーを、遅延(最初のトークンまでの時間)、出力速度(1秒あたりの出力トークン数)、料金などの性能指標で分析します。ベンチマーク対象のAPIプロバイダーにはが含まれます。
最速
出力速度
プロバイダー合計:0社
最短の遅延
最初のトークンまでの時間
プロバイダー合計:0社
最安料金
ブレンド料金(100万トークンあたり)
プロバイダー合計:0社
現在、Llama 3.2 90B (Vision)を利用できるAPIプロバイダーはありません。
このモデルではプロバイダーのベンチマークを利用できません。
モデルの詳細と他のモデルとの知能比較は、Llama 3.2 Instruct 90B (Vision)のモデルページをご覧ください。
ハイライト
更新:本番環境のユースケースをより適切に反映するため、性能ベンチマークのデフォルトワークロードを入力10kトークンに変更しました。上部で別のワークロードを引き続き選択できます。
料金
Pricing: Cache Hit, Input, and Output
Price (USD per M Tokens) · Lower is better · 10,000 input tokens
No data available
Pricing: Blended Price
Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended) · Lower is better
No data available
Output Speed vs. Price
Blended at 7:2:1 (cache-input-output) · Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
No data available
速度
出力速度(1秒あたりのトークン数)で測定
Output Speed: Llama 3.2 Instruct 90B (Vision)
Output speed: output tokens per second · 10,000 input tokens
No data available
Latency vs. Output Speed
Latency: seconds to first token received · Output speed: output tokens per second · 10,000 input tokens
Most attractive quadrant
No data available
遅延
最初のトークンまでの時間(秒)で測定
最初のトークンまでの時間:Llama 3.2 90B (Vision)のプロバイダー
Seconds to first token received · Lower is better · 10,000 input tokens
No data available
エンドツーエンド応答時間
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
エンドツーエンド応答時間:Llama 3.2 90B (Vision)のプロバイダー
Seconds to output 500 tokens, including reasoning model 'thinking' time · Lower is better · 10,000 input tokens
No data available
主要な比較指標とAPI機能
| No results. | |||||||||
よくある質問
Llama 3.2 Instruct 90B (Vision)のプロバイダーに関するよくある質問
現在、ベンチマーク対象のAPIプロバイダーでLlama 3.2 Instruct 90B (Vision)を利用することはできません。オープンウェイトモデルのため、セルフホストできます。