Zyphra:モデルの知能、性能、料金

この分析は、ユースケースに最適なZyphra提供モデルを選ぶための参考情報です。
最高の知能
UpdatedIntelligence Index
モデル合計:3件
最速
出力速度
モデル合計:3件
最安料金
ブレンド料金(100万トークンあたり)
モデル合計:3件
Zyphraは3モデルを提供しており、知能、性能、料金の特性はそれぞれ異なります。 以下でモデル間の主要指標を比較します。
- Zyphraで知能が上位のモデルはMiMo-V2.5-Pro(26)、MiMo-V2.5(25)、MiMo-V2.5-Pro (Non-reasoning)(18)です。
- 出力速度が最も速いモデルはMiMo-V2.5(91 t/s)、MiMo-V2.5-Pro (Non-reasoning)(82 t/s)、MiMo-V2.5-Pro(81 t/s)です。
- 遅延では、MiMo-V2.5-Pro (Non-reasoning)(0.80秒)、MiMo-V2.5(22.68秒)、MiMo-V2.5-Pro(25.45秒)の最初の回答トークンまでの時間が最短です。
- 料金では、MiMo-V2.5($0.20)、MiMo-V2.5-Pro($0.51)、MiMo-V2.5-Pro (Non-reasoning)($0.51)の100万トークンあたりのブレンド料金が最安です。
- Zyphraで最大のコンテキストウィンドウに対応するモデルはMiMo-V2.5-Pro(1M)、MiMo-V2.5(1M)、MiMo-V2.5-Pro (Non-reasoning)(1M)です。
- MiMo-V2.5は出力が最も速く、料金も最も優れているため、スループットと費用を重視するアプリケーションに適しています。最高品質が必要なタスクでは、MiMo-V2.5-Proが知能で首位です。
ハイライト
知能評価
Artificial Analysis Intelligence Index
Intelligence Evaluations
Agentic knowledge work, (Elo-500)/2000
Agentic real-world work tasks, (Elo-500)/2000
Agentic SaaS workflows
Agentic coding & terminal use
Coding
Reasoning & knowledge
Professional document reasoning, All-pass
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Legal agentic work, criterion pass rate
Agentic business operations
Quantitative analysis on spreadsheets & documents
Agentic tool use
Kubernetes incident root-cause analysis
Visual reasoning
Medical long context reasoning
Intelligence Index vs. Price
コンテキストウィンドウ
Context Window
料金
Intelligence Index vs. Price
性能の概要
Output Speed vs. Price
速度
出力速度(1秒あたりのトークン数)で測定
Output Speed
遅延
最初のトークンまでの時間(秒)で測定
Latency: Time To First Answer Token
エンドツーエンド応答時間
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
End-to-End Response Time vs. Price
主要な定義
よくある質問
Zyphraに関するよくある質問
Zyphraが提供し、当社が追跡しているモデルは3モデルです:MiMo-V2.5-Pro、MiMo-V2.5、MiMo-V2.5-Pro (Non-reasoning)。
Zyphraで利用できるモデルのうち、知能が最も高いのはIntelligence Indexスコア26のMiMo-V2.5-Proです。
Zyphraで出力速度が最も速いモデルは、毎秒91.2トークンのMiMo-V2.5です。
Zyphraで最初の回答トークンまでの時間が最短のモデルは、0.80秒のMiMo-V2.5-Pro (Non-reasoning)です。遅延が短いほど、最初の応答が速くなります。
Zyphraでブレンド料金が最も安いモデルは、100万トークンあたり$0.20のMiMo-V2.5です(キャッシュヒット/入力/出力を7:2:1とした場合)。
Zyphraのモデル間では料金に最大3倍の差があり、MiMo-V2.5の100万トークンあたり$0.20から、MiMo-V2.5-Pro (Non-reasoning)の$0.51までとなっています。
はい。ZyphraはOpenAI互換APIを提供しているため、OpenAIからの切り替えや既存のOpenAI SDK連携の利用が容易です。
Zyphraの3モデル中2モデルが、構造化出力のJSONモードに対応しています。
はい。Zyphraの全3モデルが関数呼び出し(ツール利用)に対応しています。
はい。Zyphraは2推論モデルを提供しています:MiMo-V2.5-Pro、MiMo-V2.5。推論モデルは回答前に拡張思考を行い、複雑な問題に取り組みます。
はい。Zyphraの全3モデルがオープンウェイトです。
はい。インフラストラクチャの変更、負荷分散、アップデートにより、プロバイダーの性能は時間とともに変化する場合があります。すべてのプロバイダーを継続的にベンチマークし、「推移」グラフに過去の性能傾向を表示しています。
Zyphraのモデルを選ぶ際は、知能(品質を重視するタスク)、出力速度(高スループットが必要なタスク)、遅延(最初の応答の速さが必要な対話型アプリケーション)、料金(費用を重視するワークロード)、コンテキストウィンドウの規模、JSONモード、関数呼び出しへの対応などを検討してください。