Microsoft 模型:智能、性能与价格
Artificial Analysis 已对 Microsoft 的 4 个模型进行基准测试。下方对这些模型的关键指标进行了比较。
- 智能方面,Microsoft 得分最高的模型是 Phi-4 Mini Instruct(6,估算值)。
- 输出速度方面,最快的模型是 Phi-4 Mini Instruct(45 token/秒)。
- 延迟方面,Phi-4 Multimodal Instruct(0.79 秒) 的首个答案 token 等待时间最短。
智能
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
成本
Cost per Intelligence Index Task
速度与延迟
Output Speed
能力得分
能力指数
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Microsoft 的所有发布
更多详情
权重 | 服务商基准测试 | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| Phi-4 Mini Instruct | 6 | 3.8B | 128k | US$0.0 | 45 | ||||
| Phi-4 | 6 | 14B | 16k | US$0.2 | 40 | ||||
| Phi-3 Mini Instruct 3.8B | 6 | 3.8B | 4k | - | - | - | |||
| Phi-4 Multimodal Instruct | 6 | 5.6B | 128k | US$0.0 | 17 |