Institute of Foundation Modelsのモデル:知能・性能・料金
Artificial Analysis Intelligence Index
* 推定値
Intelligence Indexのタスクあたり
Artificial AnalysisはInstitute of Foundation Modelsの9個のモデルをベンチマークしています。以下では、これらのモデルの主要指標を比較します。
- 知能が最も高いInstitute of Foundation ModelsのモデルはK2 Horizon 375B A23B(31)です。
知能
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
費用
Cost per Intelligence Index Task
速度と遅延
Output Speed
能力スコア
能力指数
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Institute of Foundation Modelsのすべてのリリース
詳細
ウェイト | プロバイダーのベンチマーク | ||||||||
|---|---|---|---|---|---|---|---|---|---|
| K2 Horizon 375B A23B | 31 | 375B 推論時に23Bが有効 | 524k | - | - | - | |||
| K2 Horizon MoVA 36B A4B | 25 | 36B 推論時に4Bが有効 | 524k | - | - | - | |||
| K2 Horizon 7B | 21 | 7B | 524k | - | - | - | |||
| K2 Horizon 3.7B | 16 | 3.7B | 524k | - | - | - | |||
| K2 Think V2 | 11 | 70B | 262k | - | - | - | |||
| K2-V2 (high) | 10 | 70B | 512k | - | - | - | |||
| K2-V2 (medium) | 9 | 70B | 512k | - | - | - | |||
| K2-V2 (low) | 7 | 70B | 512k | - | - | - | |||
| K2 Horizon 0.9B | 3 | 0.9B | 131k | - | - | - |