IBM-Modelle: Intelligenz, Leistung und Preis

Artificial Analysis hat 13 Modelle von IBM gebenchmarkt. Unten werden die wichtigsten Kennzahlen dieser Modelle verglichen.

  • Bei der Intelligenz ist Granite 4.2 30B mit 15 (geschätzt) das führende Modell von IBM.
  • Bei der Ausgabegeschwindigkeit ist Granite 4.2 3B mit 216 Tokens/s das schnellste Modell.
  • Bei der Latenz bietet Granite 4.2 3B mit 9,77 s die kürzeste Zeit bis zum ersten Antwort-Token.
  • Beim Preis bietet Granite 4.2 3B mit $0.01 die niedrigsten Kosten pro Aufgabe. Die Preise unterscheiden sich zwischen den Modellen um einen Faktor von bis zu 3,9.

Intelligenz

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
Pareto line

Kosten

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Geschwindigkeit und Latenz

Output Speed

Output tokens per second · Higher is better

Fähigkeitswerte

Fähigkeitsindizes

Misst die Leistung von Modellen bei bestimmten Fähigkeiten und in bestimmten Branchen
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

No data available
Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better

Alle Veröffentlichungen von IBM

Weitere Details

Gewichte
Anbieter-Benchmarks
Granite 4.2 30B
Logo von IBMIBM
15
30B
131k
0,1 $
77
DeepInfra
Granite 4.2 8B
Logo von IBMIBM
11
8B
131k
0,0 $
76
CoreWeaveDeepInfra
Granite 4.2 3B
Logo von IBMIBM
9
3B
131k
0,0 $
216
DeepInfra
Granite 4.1 30B
Logo von IBMIBM
7
30B
131k
-
-
-
Granite 4.1 8B
Logo von IBMIBM
7
8B
131k
0,1 $
126
CoreWeave
Granite 4.0 H Small
Logo von IBMIBM
6
32B
9B während der Inferenz aktiv
128k
0,1 $
15
Replicate
Granite 4.1 3B
Logo von IBMIBM
6
3B
131k
-
-
-
Granite 4.0 H 1B
Logo von IBMIBM
5
1.5B
128k
-
-
-
Granite 4.0 Micro
Logo von IBMIBM
5
3B
128k
-
-
-
Granite 4.0 1B
Logo von IBMIBM
5
1.6B
128k
-
-
-
Granite 3.3 8B (Non-reasoning)
Logo von IBMIBM
5
8.2B
128k
0,1 $
15
Replicate
Granite 4.0 350M
Logo von IBMIBM
5
0.3B
33k
-
-
-
Granite 4.0 H 350M
Logo von IBMIBM
5
0.3B
33k
-
-
-