Microsoftのモデル:知能・性能・料金
Artificial AnalysisはMicrosoftの4個のモデルをベンチマークしています。以下では、これらのモデルの主要指標を比較します。
- 知能が最も高いMicrosoftのモデルはPhi-4 Mini Instruct(6、推定)です。
- 出力速度が最も速いモデルはPhi-4 Mini Instruct(45トークン/秒)です。
- 最初の回答トークンまでの時間が最も短いモデルはPhi-4 Multimodal Instruct(0.79秒)です。
知能
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
費用
Cost per Intelligence Index Task
速度と遅延
Output Speed
能力スコア
能力指数
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Microsoftのすべてのリリース
詳細
ウェイト | プロバイダーのベンチマーク | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| Phi-4 Mini Instruct | 6 | 3.8B | 128k | $0.0 | 45 | ||||
| Phi-4 | 6 | 14B | 16k | $0.2 | 40 | ||||
| Phi-3 Mini Instruct 3.8B | 6 | 3.8B | 4k | - | - | - | |||
| Phi-4 Multimodal Instruct | 6 | 5.6B | 128k | $0.0 | 17 |