Modelos de Microsoft: inteligência, desempenho e preço
A Artificial Analysis avaliou 4 modelos de Microsoft. Abaixo, comparamos as principais métricas desses modelos.
- Em inteligência, o principal modelo de Microsoft é Phi-4 Mini Instruct, com 6 (estimativa).
- Em velocidade de saída, o modelo mais rápido é Phi-4 Mini Instruct, com 45 tokens/s.
- Em latência, Phi-4 Multimodal Instruct, com 0,79 s oferece o menor tempo até o primeiro token de resposta.
Inteligência
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Custo
Cost per Intelligence Index Task
Velocidade e latência
Output Speed
Pontuações de capacidades
Índices de capacidades
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Todos os lançamentos de Microsoft
Mais detalhes
Pesos | Benchmarks de provedores | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| Phi-4 Mini Instruct | 6 | 3.8B | 128k | US$ 0,0 | 45 | ||||
| Phi-4 | 6 | 14B | 16k | US$ 0,2 | 40 | ||||
| Phi-3 Mini Instruct 3.8B | 6 | 3.8B | 4k | - | - | - | |||
| Phi-4 Multimodal Instruct | 6 | 5.6B | 128k | US$ 0,0 | 17 |