Allen Institute for AI 모델: 지능, 성능 및 가격
Artificial Analysis Intelligence Index
* 추정치
Intelligence Index 작업당
Artificial Analysis는 Allen Institute for AI의 모델 10개를 벤치마크했습니다. 아래에서 이 모델들의 주요 지표를 비교합니다.
- 지능 부문에서 Allen Institute for AI의 최고 모델은 Olmo 3.1 32B Think(7, 추정)입니다.
지능
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
비용
Cost per Intelligence Index Task
속도 및 지연 시간
Output Speed
역량 점수
역량 지수
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Allen Institute for AI의 모든 릴리스
상세 정보
웨이트 | 제공업체 벤치마크 | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| Llama 3.1 Tulu3 405B | 7 | 405B | 128k | - | - | - | |||
| Olmo 3.1 32B Think | 7 | 32.2B | 66k | US$0.0 | - | ||||
| Olmo 3.1 32B Instruct | 6 | 32.2B | 66k | - | - | - | |||
| Olmo 3 32B Think | 6 | 32.2B | 66k | - | - | - | |||
| OLMo 2 32B | 6 | 32.2B | 4k | - | - | - | |||
| Olmo 3 7B Think | 6 | 7B | 66k | - | - | - | |||
| OLMo 2 7B | 6 | 7.3B | 4k | - | - | - | |||
| Molmo 7B-D | 6 | 8.0B | 4k | - | - | - | |||
| Olmo 3 7B Instruct | 5 | 7B | 66k | US$0.1 | - | ||||
| Molmo2-8B | 5 | 8.7B | 37k | - | - | - |