LLM Leaderboard - Comparison of AI models from OpenAI, Anthropic, Google, SpaceXAI & others
Comparison and ranking the performance of over 250 AI models (LLMs) across key metrics including intelligence, price, performance and speed (output speed - tokens per second & latency - TTFT), context window & others.
For more details including relating to our methodology, see our FAQs.
Intelligence
UpdatedClaude Fable 5.1 (max with fallback) and Claude Fable 5.1 (xhigh with fallback) are the highest intelligence models, followed by GPT-6 Astra (max) and GPT-6 Astra (xhigh).
Output Speed
Celeris-1 and Mercury 2 are the fastest models, followed by Gemini 2.5 Flash-Lite and Gemini 3.5 Flash-Lite.
Latency
Gemini 2.5 Flash-Lite (Non-reasoning) and North Mini Code are the lowest latency models, followed by Granite 4.2 3B and Command A+.
Cost per Task
Granite 4.2 3B and GPT-5.6 Luna (low) have the lowest cost per task, followed by GPT-5.6 Luna (Non-reasoning) and gpt-oss-20b (high).
Context Window
Llama 4 Scout and Grok 4.20 0309 support the largest context windows, followed by Gemini 1.5 Pro (May) and Grok 4.1 Fast.
Further Analysis | ||||||||
|---|---|---|---|---|---|---|---|---|
Claude Fable 5.1 (max with fallback) | 1M | 53 | $7.63 | 70 | 262.21 | 269.39 | ||
Claude Fable 5.1 (xhigh with fallback) | 1M | 53 | $5.98 | 61 | 157.09 | 165.27 | ||
GPT-6 Astra (max) | 1M | 53 | $3.26 | 59 | 260.42 | 268.84 | ||
GPT-6 Astra (xhigh) | 1M | 52 | $2.31 | 54 | 154.72 | 164.00 | ||
Claude Fable 5.1 (high with fallback) | 1M | 51 | $3.91 | 55 | 14.70 | 23.78 | ||
GPT-6 Astra (high) | 1M | 51 | $1.73 | 52 | 36.07 | 45.62 | ||
Claude Opus 5 (max) | 1M | 51 | $5.86 | 55 | 56.84 | 65.88 | ||
Claude Opus 5 (xhigh) | 1M | 50 | $4.88 | 54 | 21.36 | 30.68 | ||
GPT-6 Astra (medium) | 1M | 50 | $1.54 | 53 | 4.88 | 14.32 | ||
Claude Fable 5.1 (medium with fallback) | 1M | 49 | $2.98 | 55 | 6.45 | 15.55 | ||
Claude Opus 5 (high) | 1M | 48 | $3.61 | 54 | 16.69 | 25.91 | ||
Muse Spark 1.3 (max) | 1M | 48 | $1.60 | 224 | 26.72 | 37.88 | ||
GPT-5.6 Sol (max) | 1M | 47 | $1.99 | 61 | 129.50 | 137.64 | ||
Claude Fable 5.1 (low with fallback) | 1M | 47 | $2.37 | 54 | 6.59 | 15.88 | ||
GPT-6 Astra (low) | 1M | 46 | $0.82 | 55 | 2.56 | 11.71 | ||
Qwen3.8 Max (0902) | 984k | 45 | $5.41 | 37 | 2.72 | 69.94 | ||
Muse Spark 1.3 (xhigh) | 1M | 45 | $1.37 | 250 | 23.58 | 33.60 | ||
Claude Opus 5 (medium) | 1M | 45 | $2.19 | 55 | 3.84 | 13.01 | ||
GLM-5.3 (max) | 1M | 45 | $2.01 | 72 | 2.99 | 37.66 | ||
Grok 4.6 (high) | 500k | 44 | $1.86 | 60 | 41.52 | 49.88 | ||
Grok 4.6 (xhigh) | 500k | 44 | $2.32 | 60 | 36.05 | 44.35 | ||
GPT-5.6 Sol (xhigh) | 1M | 44 | $1.18 | 65 | 38.70 | 46.36 | ||
Step 5 Preview | 1M | 44 | $0.71 | 100 | 2.96 | 28.02 | ||
Kimi K3 (max) | 1.05M | 44 | $2.00 | 39 | 3.99 | 67.60 | ||
Grok 4.6 (medium) | 500k | 43 | $1.50 | 59 | 34.07 | 42.57 | ||
GPT-5.6 Sol (high) | 1M | 42 | $0.81 | 63 | 9.84 | 17.74 | ||
GPT-5.6 Terra (max) | 1M | 42 | $1.40 | 89 | 211.05 | 216.64 | ||
GLM-5.3-Flash | 1M | 42 | $0.25 | 95 | 2.53 | 28.80 | ||
Gemini 3.8 Flash (high) | 1M | 41 | $1.24 | 305 | 16.44 | 18.07 | ||
Qwen3.8 2.4T A95B | 984k | 40 | $2.16 | 38 | 2.69 | 68.00 | ||
Qwen3.8-Flash-Next | 256k | 40 | $0.37 | 57 | 2.98 | 46.98 | ||
Gemini 3.8 Flash (medium) | 1M | 40 | $0.93 | -- | -- | -- | ||
DeepSeek V4.1 Flash (max) | 1M | 39 | $0.27 | 207 | 1.15 | 13.20 | ||
Claude Opus 5 (low) | 1M | 39 | $1.10 | 56 | 3.03 | 12.04 | ||
GPT-5.6 Sol (medium) | 1M | 39 | $0.50 | 60 | 5.09 | 13.36 | ||
Claude Sonnet 5 (max) | 1M | 38 | $5.09 | 72 | 131.49 | 138.40 | ||
GPT-5.6 Terra (xhigh) | 1M | 38 | $0.63 | 81 | 18.27 | 24.42 | ||
GPT-5.6 Luna (max) | 1M | 37 | $0.18 | 162 | 123.28 | 126.37 | ||
DeepSeek V4 Pro 0813 (max) | 1M | 36 | $0.67 | 89 | 1.69 | 29.84 | ||
Agnes 3.0 Flash | 1M | 36* | -- | -- | -- | -- | ||
Agnes 2.5 Pro Beta | 1M | 35* | -- | -- | -- | -- | ||
Grok 4.6 (low) | 500k | 35 | $0.48 | 55 | 7.65 | 16.68 | ||
DeepSeek V4 Flash Vision (max) | 1M | 35 | $0.31 | 212 | 1.04 | 12.83 | ||
GPT-5.6 Luna (xhigh) | 1M | 35 | $0.09 | 143 | 34.58 | 38.09 | ||
Kimi K3 (low) | 1.05M | 34* | -- | 37 | 4.16 | 72.00 | ||
Claude Sonnet 5 (xhigh) | 1M | 34 | $2.87 | 70 | 32.46 | 39.57 | ||
GPT-5.6 Terra (high) | 1M | 34 | $0.34 | 89 | 2.69 | 8.33 | ||
Qwen3.8 27B (xhigh) | 256k | 34 | $0.82 | 45 | 3.82 | 59.85 | ||
Motif 3 | 262k | 34* | -- | -- | -- | -- | ||
GPT-5.6 Sol (low) | 1M | 33 | $0.26 | 65 | 2.87 | 10.53 | ||
Gemini 3.8 Flash (low) | 1M | 33 | -- | -- | -- | -- | ||
GPT-5.3 Codex (xhigh) | 400k | 33* | -- | 140 | 54.30 | 57.86 | ||
Motif 3 (Beta) | 262k | 32* | -- | -- | -- | -- | ||
GPT-5.6 Luna (high) | 1M | 32 | $0.04 | 140 | 8.15 | 11.72 | ||
Claude Sonnet 5 (high) | 1M | 32 | $1.79 | 65 | 5.41 | 13.08 | ||
K2 Horizon 375B A23B | 524k | 31 | -- | -- | -- | -- | ||
Apodex 1.1 | 256k | 30* | -- | -- | -- | -- | ||
GPT-5.6 Terra (medium) | 1M | 30 | $0.18 | 84 | 2.01 | 7.94 | ||
Gemini 3.1 Pro Preview | 1M | 30 | $0.67 | 119 | 35.72 | 39.92 | ||
MiniMax-M3 | 1M | 29 | $0.51 | 210 | 1.01 | 12.89 | ||
GPT-5.6 Sol (Non-reasoning) | 1M | 28* | -- | 60 | 1.17 | 9.52 | ||
Nex-N2-Pro | 262k | 28* | -- | -- | -- | -- | ||
Solar Pro 4 | 512k | 28* | -- | 71 | 2.03 | 37.01 | ||
Claude Sonnet 5 (medium) | 1M | 28 | $1.00 | 63 | 2.39 | 10.29 | ||
Inkling Small | 1M | 28* | -- | 171 | 2.07 | 16.67 | ||
Qwen3.8 27B (medium) | 256k | 28 | $0.90 | 48 | 3.87 | 56.36 | ||
GPT-5.6 Terra (low) | 1M | 27 | $0.14 | 79 | 2.16 | 8.47 | ||
JT-4.1 Flash 236B A21B | 256k | 27* | -- | -- | -- | -- | ||
Agnes 2.5 Pro Alpha | 1M | 27* | -- | 186 | 3.97 | 17.40 | ||
Quasar 438B (max) | 1M | 27 | $2.02 | 175 | 1.11 | 15.36 | ||
Qwen3.8 27B (low) | 256k | 26 | $0.83 | 54 | 3.89 | 50.02 | ||
GPT-5.5 Instant (June 2026) | 400k | 26 | $0.69 | 131 | 1.02 | 20.11 | ||
MiMo-V2.5-Pro | 1M | 26 | $0.05 | 52 | 4.06 | 52.11 | ||
Kimi K2.7 Code | 256k | 26 | $0.54 | 55 | 2.92 | 52.24 | ||
K2 Horizon MoVA 36B A4B | 524k | 25 | -- | -- | -- | -- | ||
Hy3 | 256k | 25 | $0.07 | 81 | 2.68 | 33.47 | ||
MiMo-V2.5 | 1M | 25* | -- | 31 | 6.29 | 86.02 | ||
Qwen3.7 Plus | 1M | 25 | $0.33 | 67 | 2.22 | 39.26 | ||
GPT-5.6 Luna (medium) | 1M | 25 | $0.02 | 136 | 2.48 | 6.17 | ||
Inkling | 1M | 25 | $0.61 | 84 | 2.57 | 32.25 | ||
Ling 3.0 Flash | 262k | 25* | -- | 320 | 2.54 | 10.35 | ||
Solar Open2 250B | 1.05M | 25* | -- | -- | -- | -- | ||
Ling-3.0-flash-VL | 262k | 25 | -- | 144 | 2.30 | 19.71 | ||
Claude Sonnet 5 (low) | 1M | 24 | $0.51 | 59 | 1.59 | 10.02 | ||
Claude Sonnet 5 (Non-reasoning) | 1M | 23* | -- | 64 | 1.30 | 9.08 | ||
Nemotron 3 Ultra | 262k | 23 | $0.58 | 165 | 2.45 | 19.31 | ||
A.X-K2 | 262k | 23* | -- | -- | -- | -- | ||
Ling-3.0-flash-Fin | 262k | 23 | -- | 159 | 2.97 | 18.67 | ||
MiMo-V2-Flash (Feb 2026) | 256k | 22* | -- | -- | -- | -- | ||
Qwen3.8 27B | 256k | 22* | -- | 46 | 3.91 | 14.80 | ||
Gemini 3.5 Flash-Lite | 1M | 22 | $0.12 | 398 | 9.78 | 11.04 | ||
G9v3-39A5B | 131k | 22* | -- | -- | -- | -- | ||
KAT-Coder-Pro V2 | 256k | 22* | -- | -- | -- | -- | ||
Qwen3.5 397B A17B (Non-reasoning) | 262k | 21* | -- | 87 | 2.17 | 7.92 | ||
GPT-5.6 Luna (low) | 1M | 21 | $0.01 | 135 | 1.97 | 5.67 | ||
GPT-5.6 Terra (Non-reasoning) | 1M | 21 | $0.14 | 84 | 0.97 | 6.90 | ||
K2 Horizon 7B | 524k | 21 | -- | -- | -- | -- | ||
Qwen3.5 Omni Plus | 256k | 20* | -- | 94 | 2.09 | 7.39 | ||
o3 | 200k | 20* | -- | 130 | 5.44 | 9.28 | ||
K-EXAONE 2.0 | 262k | 20* | -- | -- | -- | -- | ||
Step 3.7 Flash | 262k | 19* | -- | 168 | 2.55 | 17.40 | ||
LongCat 2.0 | 1M | 19 | $0.06 | -- | -- | -- | ||
Gemma 4 31B | 256k | 19* | -- | 35 | 0.98 | 64.46 | ||
JT-35B-Flash | 256k | 19* | -- | -- | -- | -- | ||
Qwen3.5 397B A17B | 262k | 18 | $0.47 | 86 | 2.12 | 44.90 | ||
MiMo-V2.5-Pro (Non-reasoning) | 1M | 18* | -- | 50 | 4.17 | 14.23 | ||
Qwen3.6 35B A3B | 262k | 18 | $0.48 | 109 | 2.11 | 56.22 | ||
Qwen3.5 122B A10B (Non-reasoning) | 262k | 18* | -- | 144 | 2.34 | 5.82 | ||
Muse Glimmer (high) | 131k | 17 | $0.06 | 94 | 0.92 | 27.53 | ||
Doubao Seed Code | 256k | 17* | -- | -- | -- | -- | ||
Claude 4.5 Haiku | 200k | 17 | $0.21 | 92 | 17.18 | 22.60 | ||
Gemma 4 26B A4B | 256k | 17* | -- | -- | -- | -- | ||
Ring-2.6-1T | 262k | 17 | $0.29 | 131 | 3.76 | 22.92 | ||
K2 Horizon 3.7B | 524k | 16 | -- | -- | -- | -- | ||
Qwen3.5 122B A10B | 262k | 16 | $0.32 | 129 | 2.32 | 21.70 | ||
GPT-5.6 Luna (Non-reasoning) | 1M | 16 | $0.01 | 133 | 0.83 | 4.60 | ||
Claude 4.5 Haiku (Non-reasoning) | 200k | 15* | -- | 86 | 0.69 | 6.48 | ||
Ling 3.0 Tiny | 262k | 15* | -- | 156 | 4.26 | 20.33 | ||
Qwen3.6 35B A3B (Non-reasoning) | 262k | 15* | -- | 111 | 2.13 | 6.63 | ||
Granite 4.2 30B | 131k | 15* | -- | 75 | 0.84 | 34.17 | ||
ERNIE 5.0 Thinking Preview | 128k | 14* | -- | -- | -- | -- | ||
Mistral Medium 3.5 | 256k | 14 | $0.44 | 141 | 2.29 | 20.05 | ||
Gemma 4 12B | 256k | 14* | -- | 110 | 2.45 | 25.24 | ||
Nova 2.0 Pro Preview (medium) | 256k | 14* | -- | 116 | 14.66 | 36.28 | ||
Gemma 4 31B (Non-reasoning) | 256k | 14* | -- | 42 | 2.10 | 14.14 | ||
Mercury 2 | 128k | 14* | -- | 414 | 6.20 | 7.41 | ||
Qwen3.5 9B | 262k | 14* | -- | 71 | 1.11 | 36.56 | ||
Nova 2.0 Omni (medium) | 1M | 14* | -- | -- | -- | -- | ||
Apriel-v1.6-15B-Thinker | 128k | 13* | -- | -- | -- | -- | ||
Nova 2.0 Lite (high) | 1M | 13* | -- | 159 | 17.83 | 33.54 | ||
Qwen3.5 9B (Non-reasoning) | 262k | 13* | -- | 92 | 0.72 | 6.17 | ||
EXAONE 4.5 33B | 262k | 13* | -- | -- | -- | -- | ||
Command A+ | 192k | 13 | $0.00 | 241 | 0.42 | 10.80 | ||
Gemma 4 26B A4B (Non-reasoning) | 256k | 13* | -- | 67 | 1.44 | 8.94 | ||
Qwen3.5 4B | 262k | 13* | -- | 19 | 0.85 | 134.87 | ||
Nemotron 3.5 Lightning | 1M | 13 | $0.09 | 297 | 0.59 | 9.01 | ||
Nemotron 3 Super | 1M | 13 | $1.06 | 187 | 1.79 | 15.17 | ||
Nova 2.0 Pro Preview (low) | 256k | 13* | -- | 115 | 10.41 | 32.24 | ||
Nova 2.0 Lite (medium) | 1M | 12* | -- | 162 | 15.42 | 30.90 | ||
Qwen3.5 Omni Flash | 256k | 12* | -- | 229 | 1.83 | 4.01 | ||
MiniCPM5-2B | 131k | 12 | -- | -- | -- | -- | ||
JT-MINI | 128k | 12* | -- | -- | -- | -- | ||
Magistral Medium 1.2 | 128k | 12* | -- | -- | -- | -- | ||
Nova 2.0 Lite (low) | 1M | 12* | -- | 162 | 9.49 | 24.91 | ||
HyperNova 60B 2605 (high) | 131k | 12* | -- | -- | -- | -- | ||
Nemotron Cascade 2 30B A3B | 1M | 12* | -- | -- | -- | -- | ||
gpt-oss-120b (high) | 131k | 12 | $0.11 | 185 | 0.84 | 14.33 | ||
K2 Think V2 | 262k | 11* | -- | -- | -- | -- | ||
LongCat Flash Lite | 256k | 11* | -- | -- | -- | -- | ||
HyperCLOVA X SEED Think (32B) | 128k | 11* | -- | -- | -- | -- | ||
Mistral Small 4 | 256k | 11 | $0.05 | 169 | 0.80 | 15.63 | ||
Qwen3 Next 80B A3B (Reasoning) | 262k | 11* | -- | 180 | 2.25 | 16.17 | ||
Nova 2.0 Omni (low) | 1M | 11* | -- | -- | -- | -- | ||
Granite 4.2 8B | 131k | 11 | $0.02 | 84 | 0.87 | 30.79 | ||
Mi:dm K 2.5 Pro | 128k | 11* | -- | -- | -- | -- | ||
G9v3-3B | 131k | 11* | -- | -- | -- | -- | ||
Trinity Large Thinking | 512k | 11 | $0.35 | 303 | 1.16 | 9.41 | ||
Qwen3.5 4B (Non-reasoning) | 262k | 11* | -- | 18 | 0.75 | 28.54 | ||
INTELLECT-3 | 131k | 11* | -- | -- | -- | -- | ||
Solar Open 100B | 128k | 10* | -- | -- | -- | -- | ||
Nemotron 3 Nano Omni 30B A3B | 256k | 10* | -- | -- | -- | -- | ||
gpt-oss-120b (low) | 131k | 10* | -- | 196 | 0.84 | 13.60 | ||
Llama 4 Maverick | 1M | 10* | -- | 110 | 0.85 | 5.38 | ||
Nova 2.0 Pro Preview (Non-reasoning) | 256k | 10* | -- | 100 | 1.04 | 6.07 | ||
gpt-oss-20b (low) | 131k | 10* | -- | 201 | 0.98 | 13.39 | ||
North Mini Code | 256k | 10 | $0.00 | 83 | 0.38 | 30.60 | ||
K2-V2 (high) | 512k | 10* | -- | -- | -- | -- | ||
Qwen3 Next 80B A3B | 262k | 10* | -- | 178 | 2.23 | 5.04 | ||
DiffusionGemma 26B A4B | 256k | 10* | -- | -- | -- | -- | ||
Gemma 4 12B (Non-reasoning) | 262k | 9* | -- | 110 | 2.48 | 7.04 | ||
Mistral Large 3 | 256k | 9 | $0.10 | 77 | 1.06 | 7.57 | ||
Qwen3 Coder Next | 256k | 9 | $0.55 | 112 | 1.35 | 5.83 | ||
Motif-2-12.7B | 128k | 9* | -- | -- | -- | -- | ||
Nova Premier | 1M | 9* | -- | 30 | 2.87 | 19.44 | ||
Granite 4.2 3B | 131k | 9 | $0.01 | 215 | 0.40 | 12.03 | ||
K2-V2 (medium) | 512k | 9* | -- | -- | -- | -- | ||
Llama Nemotron Super 49B v1.5 | 128k | 9* | -- | 68 | 5.51 | 42.28 | ||
Mistral Small 4 (Non-reasoning) | 256k | 9* | -- | 152 | 0.96 | 4.25 | ||
Tri-21B-Think | 32k | 9* | -- | -- | -- | -- | ||
gpt-oss-20b (high) | 131k | 9 | $0.01 | 168 | 0.79 | 15.69 | ||
Gemma 4 E4B | 128k | 9* | -- | 41 | 0.92 | 62.07 | ||
Nemotron 3 Nano | 1M | 9 | $0.02 | 124 | 1.29 | 21.38 | ||
MiniCPM5-1B | 128k | 9* | -- | -- | -- | -- | ||
Sarvam 105B (high) | 128k | 9* | -- | -- | -- | -- | ||
Nova 2.0 Lite (Non-reasoning) | 1M | 9* | -- | 149 | 1.06 | 4.41 | ||
MiniCPM5-1B (Non-reasoning) | 128k | 9* | -- | -- | -- | -- | ||
Devstral 2 | 256k | 9 | $0.00 | 150 | 2.27 | 5.61 | ||
Magistral Small 1.2 | 128k | 9* | -- | -- | -- | -- | ||
Nanbeige4.1-3B | 256k | 8* | -- | -- | -- | -- | ||
LFM2.5-2.6B | 128k | 8* | -- | -- | -- | -- | ||
EXAONE 4.0 32B | 131k | 8* | -- | -- | -- | -- | ||
Nova 2.0 Omni (Non-reasoning) | 1M | 8* | -- | -- | -- | -- | ||
Llama 4 Scout | 10M | 8* | -- | 100 | 0.83 | 5.83 | ||
Hermes 4 70B | 128k | 8* | -- | -- | -- | -- | ||
Falcon-H1R-7B | 256k | 8* | -- | -- | -- | -- | ||
Gemma 4 E2B | 128k | 8* | -- | -- | -- | -- | ||
Qwen3 Omni 30B A3B (Reasoning) | 65.5k | 8* | -- | 101 | 2.00 | 26.83 | ||
Step3 VL 10B | 65.5k | 8* | -- | -- | -- | -- | ||
Llama 3.3 70B | 128k | 8* | -- | 86 | 1.59 | 7.41 | ||
Devstral Small 2 | 256k | 8 | $0.00 | 140 | 2.33 | 5.89 | ||
Llama Nemotron Ultra | 128k | 8* | -- | -- | -- | -- | ||
ERNIE 4.5 300B A47B | 131k | 8* | -- | -- | -- | -- | ||
Hermes 4 405B | 128k | 7* | -- | 44 | 2.37 | 58.81 | ||
NVIDIA Nemotron Nano 12B v2 VL | 128k | 7* | -- | 68 | 5.50 | 42.45 | ||
Gemma 4 E4B (Non-reasoning) | 128k | 7* | -- | 44 | 0.75 | 12.09 | ||
NVIDIA Nemotron Nano 9B V2 | 131k | 7* | -- | 115 | 4.96 | 26.77 | ||
Hermes 4 405B (Non-reasoning) | 128k | 7* | -- | 45 | 2.32 | 13.36 | ||
Nemotron 3 Nano 4B | 262k | 7* | -- | -- | -- | -- | ||
Llama Nemotron Super 49B v1.5 (Non-reasoning) | 128k | 7* | -- | 48 | 7.87 | 18.18 | ||
K2-V2 (low) | 512k | 7* | -- | -- | -- | -- | ||
Kimi Linear 48B A3B Instruct | 1M | 7* | -- | -- | -- | -- | ||
Llama 3.1 405B | 128k | 7* | -- | -- | -- | -- | ||
LFM2.5-8B-A1B | 32.8k | 7* | -- | -- | -- | -- | ||
Ring-flash-2.0 | 128k | 7* | -- | -- | -- | -- | ||
Olmo 3.1 32B Think | 65.5k | 7* | -- | -- | -- | -- | ||
Command A | 256k | 7* | -- | 60 | 1.74 | 10.03 | ||
Qwen3.5 2B | 262k | 7* | -- | -- | -- | -- | ||
Llama 3.1 Nemotron 70B | 128k | 7* | -- | 75 | 10.47 | 17.16 | ||
Nemotron 3 Nano (Non-reasoning) | 1M | 7* | -- | 187 | 0.95 | 3.63 | ||
NVIDIA Nemotron Nano 9B V2 (Non-reasoning) | 131k | 7* | -- | 156 | 2.02 | 5.24 | ||
Hermes 4 70B (Non-reasoning) | 128k | 7* | -- | -- | -- | -- | ||
Sarvam 30B (high) | 65.5k | 7* | -- | -- | -- | -- | ||
Olmo 3.1 32B Instruct | 65.5k | 6* | -- | -- | -- | -- | ||
Gemma 4 E2B (Non-reasoning) | 128k | 6* | -- | -- | -- | -- | ||
R1 1776 | 128k | 6* | -- | -- | -- | -- | ||
Llama 3.2 90B (Vision) | 128k | 6* | -- | -- | -- | -- | ||
Celeris-1 | 131k | 6 | $0.05 | 1,567 | 0.62 | 0.94 | ||
Phi-4 Mini | 128k | 6* | -- | 45 | 0.86 | 11.97 | ||
EXAONE 4.0 32B (Non-reasoning) | 131k | 6* | -- | -- | -- | -- | ||
Qwen3.5 2B (Non-reasoning) | 262k | 6* | -- | -- | -- | -- | ||
Qwen3.5 0.8B | 262k | 6* | -- | -- | -- | -- | ||
DeepHermes 3 - Mistral 24B | 32k | 6* | -- | -- | -- | -- | ||
Jamba 1.7 Large | 256k | 6* | -- | -- | -- | -- | ||
Granite 4.0 H Small | 128k | 6* | -- | 17 | 28.71 | 57.30 | ||
Ministral 3 14B | 256k | 6 | $0.14 | 87 | 0.91 | 6.65 | ||
Qwen3 Omni 30B A3B | 65.5k | 6* | -- | 96 | 1.94 | 7.16 | ||
LFM2 24B A2B | 32.8k | 6* | -- | -- | -- | -- | ||
Phi-4 | 16k | 6* | -- | 30 | 2.87 | 19.62 | ||
Nova Micro | 130k | 6* | -- | 265 | 0.85 | 2.73 | ||
NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) | 128k | 6* | -- | 182 | 1.71 | 4.46 | ||
Phi-4 Multimodal | 128k | 6* | -- | 17 | 0.86 | 29.87 | ||
MiniCPM-V 4.6 1.3B | 262k | 6* | -- | -- | -- | -- | ||
Jamba Reasoning 3B | 262k | 6* | -- | -- | -- | -- | ||
Reka Flash 3 | 128k | 6* | -- | -- | -- | -- | ||
Olmo 3 7B Think | 65.5k | 6* | -- | -- | -- | -- | ||
Molmo 7B-D | 4.1k | 6* | -- | -- | -- | -- | ||
Ministral 3 8B | 256k | 5 | $0.07 | 95 | 0.85 | 6.13 | ||
Llama 3.2 11B (Vision) | 128k | 5* | -- | 21 | 1.42 | 25.46 | ||
Qwen3.5 0.8B (Non-reasoning) | 262k | 5* | -- | -- | -- | -- | ||
Exaone 4.0 1.2B | 64k | 5* | -- | -- | -- | -- | ||
Olmo 3 7B | 65.5k | 5* | -- | -- | -- | -- | ||
Exaone 4.0 1.2B (Non-reasoning) | 64k | 5* | -- | -- | -- | -- | ||
LFM2.5-1.2B-Thinking | 32k | 5* | -- | -- | -- | -- | ||
Jamba 1.7 Mini | 258k | 5* | -- | -- | -- | -- | ||
LFM2.5-1.2B-Instruct | 32k | 5* | -- | -- | -- | -- | ||
Granite 4.0 H 1B | 128k | 5* | -- | -- | -- | -- | ||
Gemma 3 270M | 32k | 5* | -- | -- | -- | -- | ||
Apertus 70B Instruct | 65.5k | 5* | -- | -- | -- | -- | ||
Granite 4.0 Micro | 128k | 5* | -- | -- | -- | -- | ||
DeepHermes 3 - Llama-3.1 8B | 128k | 5* | -- | -- | -- | -- | ||
Molmo2-8B | 36.9k | 5* | -- | -- | -- | -- | ||
Ministral 3 3B | 256k | 5 | $0.04 | 222 | 0.62 | 2.87 | ||
LFM2.5-VL-1.6B | 32k | 5* | -- | -- | -- | -- | ||
Granite 4.0 350M | 32.8k | 5* | -- | -- | -- | -- | ||
Tiny Aya Global | 8.19k | 5* | -- | -- | -- | -- | ||
Apertus 8B Instruct | 65.5k | 5* | -- | -- | -- | -- | ||
Granite 4.0 H 350M | 32.8k | 5* | -- | -- | -- | -- | ||
K2 Horizon 0.9B | 131k | 3 | -- | -- | -- | -- | ||
EXAONE 4.5 33B (Non-reasoning) | 262k | -- | -- | -- | -- | -- | ||
Gemini 3 Deep Think | 128k | -- | -- | -- | -- | -- | ||
GPT-5.5 Pro (xhigh) | 922k | -- | -- | -- | -- | -- | ||
Cogito v2.1 | 128k | -- | -- | -- | -- | -- | ||
Key definitions
Frequently Asked Questions
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) currently ranks #1 on the Artificial Analysis LLM Leaderboard with an Intelligence Index score of 53, out of 147 models ranked.
The top models by Intelligence Index are: 1. Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) (53), 2. Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback) (53), 3. GPT-6 Astra (max) (53), 4. GPT-6 Astra (xhigh) (52), 5. Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback) (51).
Celeris-1 is the fastest at 1,566.9 tokens per second, followed by Mercury 2 (414.5 t/s) and Gemini 2.5 Flash-Lite (Reasoning) (413.0 t/s).
Granite 4.2 3B has the lowest cost per Intelligence Index task at $0.01, followed by GPT-5.6 Luna (low) ($0.01) and GPT-5.6 Luna (Non-reasoning) ($0.01).
GLM-5.3 (max) is the highest-ranked open weights model with an Intelligence Index score of 45. There are 66 open weights models out of 147 total on the leaderboard.
The top open weights models by Intelligence Index are: 1. GLM-5.3 (max) (45), 2. Kimi K3 (max) (44), 3. GLM 5.3 Flash (42).
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) leads among 130 reasoning models with an Intelligence Index score of 53. Reasoning models use extended thinking to solve complex problems before responding.
The leaderboard includes filters to narrow results by model type (reasoning vs non-reasoning), openness (open weights vs proprietary), and other criteria. You can also adjust prompt options to see how performance varies with different input lengths.
Click on any model name in the leaderboard to visit its dedicated comparison page with detailed charts covering intelligence, pricing, speed, latency, and more. You can also compare API providers for each model. View all models