모든 릴리스•

모델 10개

Allen Institute for AI 모델: 지능, 성능 및 가격

Artificial Analysis는 Allen Institute for AI의 모델 10개를 벤치마크했습니다. 아래에서 이 모델들의 주요 지표를 비교합니다.

  • 지능 부문에서 Allen Institute for AI의 최고 모델은 Olmo 3.1 32B Think(7, 추정)입니다.

지능

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
No data available

비용

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better
No data available

속도 및 지연 시간

Output Speed

Output tokens per second · Higher is better
No data available

역량 점수

역량 지수

특정 역량 및 산업에서 모델의 성능을 측정
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

No data available
Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

No data available
Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

No data available
Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

No data available
Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

No data available
Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better

No data available

Allen Institute for AI의 모든 릴리스

상세 정보

웨이트
제공업체 벤치마크
Llama 3.1 Tulu3 405B
Allen Institute for AI 로고Allen Institute for AI
7
405B
128k
-
-
-
Olmo 3.1 32B Think
Allen Institute for AI 로고Allen Institute for AI
7
32.2B
66k
US$0.0
-
Parasail
Olmo 3.1 32B Instruct
Allen Institute for AI 로고Allen Institute for AI
6
32.2B
66k
-
-
-
Olmo 3 32B Think
Allen Institute for AI 로고Allen Institute for AI
6
32.2B
66k
-
-
-
OLMo 2 32B
Allen Institute for AI 로고Allen Institute for AI
6
32.2B
4k
-
-
-
Olmo 3 7B Think
Allen Institute for AI 로고Allen Institute for AI
6
7B
66k
-
-
-
OLMo 2 7B
Allen Institute for AI 로고Allen Institute for AI
6
7.3B
4k
-
-
-
Molmo 7B-D
Allen Institute for AI 로고Allen Institute for AI
6
8.0B
4k
-
-
-
Olmo 3 7B Instruct
Allen Institute for AI 로고Allen Institute for AI
5
7B
66k
US$0.1
-
Parasail
Molmo2-8B
Allen Institute for AI 로고Allen Institute for AI
5
8.7B
37k
-
-
-