Institute of Foundation Models 모델: 지능, 성능 및 가격
Artificial Analysis Intelligence Index
* 추정치
Intelligence Index 작업당
Artificial Analysis는 Institute of Foundation Models의 모델 9개를 벤치마크했습니다. 아래에서 이 모델들의 주요 지표를 비교합니다.
- 지능 부문에서 Institute of Foundation Models의 최고 모델은 K2 Horizon 375B A23B(31)입니다.
지능
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
비용
Cost per Intelligence Index Task
속도 및 지연 시간
Output Speed
역량 점수
역량 지수
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Institute of Foundation Models의 모든 릴리스
상세 정보
웨이트 | 제공업체 벤치마크 | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| K2 Horizon 375B A23B | 31 | 375B 추론 시 23B 활성 | 524k | - | - | - | |||
| K2 Horizon MoVA 36B A4B | 25 | 36B 추론 시 4B 활성 | 524k | - | - | - | |||
| K2 Horizon 7B | 21 | 7B | 524k | - | - | - | |||
| K2 Horizon 3.7B | 16 | 3.7B | 524k | - | - | - | |||
| K2 Think V2 | 11 | 70B | 262k | - | - | - | |||
| K2-V2 (high) | 10 | 70B | 512k | - | - | - | |||
| K2-V2 (medium) | 9 | 70B | 512k | - | - | - | |||
| K2-V2 (low) | 7 | 70B | 512k | - | - | - | |||
| K2 Horizon 0.9B | 3 | 0.9B | 131k | - | - | - |