Comparisons of Large Open Source AI Models (>150B)

Open source AI models with over 150B parameters.

Models are considered open source (also commonly referred to as open weights) where their weights are accessible to download. This allows self-hosting on your own infrastructure and enables customizing the model such as through fine-tuning.

For more details including relating to our methodology, see our FAQs.

Z AI logoGLM-5.3 (max) and Kimi logoKimi K3 (max) are the highest intelligence Large open source models, defined as those with >150B parameters, followed by Z AI logoGLM-5.3-Flash & Alibaba logoQwen3.8 2.4T A95B.

Highlights

Artificial Analysis Openness Index · Higher is better
Updated
Artificial Analysis Intelligence Index · Higher is better
Trainable parameters in billions

Openness

Artificial Analysis Openness Index: Score

Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

Agentic tool use

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Size

Model Size: Total and Active Parameters

Comparison between total model parameters and parameters active during inference

Intelligence Index vs. Active Parameters

Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line

Intelligence Index vs. Total Parameters

Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Further details

Weights
Provider Benchmarks
GLM-5.3 (max)
Z AI logoZ AI
45
753B
40B active at inference time
1M
$0.9
67
ZaiSelf-hostedFireworks
+19
Kimi K3 (max)
Kimi logoKimi
44
2.8T
104B active at inference time
1M
$2.3
37
DatabricksModalTogether AI
+14
GLM 5.3 Flash
Z AI logoZ AI
42
320B
18B active at inference time
1M
$0.1
117
ZaiBasetenNovita
+15
Qwen3.8 2.4T A95B
Alibaba logoAlibaba
40
2.4T
95B active at inference time
984k
$1.2
41
FireworksTogether AIDeepInfra
+3
Qwen3.8-Flash-Next
Alibaba logoAlibaba
40
180B
6B active at inference time
256k
$0.1
53
Alibaba Cloud
DeepSeek V4.1 Flash (Reasoning, Max Effort)
DeepSeek logoDeepSeek
40
552B
16B active at inference time
1M
$0.2
220
DeepSeekNovitaFireworks
+5
DeepSeek V4 Pro 0813 (Reasoning, Max Effort)
DeepSeek logoDeepSeek
36
1.6T
49B active at inference time
1M
$0.7
93
DeepSeekGMISiliconFlow
+7
Motif 3
Motif Technologies logoMotif Technologies
34
314B
13.2B active at inference time
262k
-
-
Not available
-
K2 Horizon 375B A23B
MBZUAI Institute of Foundation Models logoMBZUAI Institute of Foundation Models
31
375B
23B active at inference time
524k
-
-
-
Kimi K3 (low)
Kimi logoKimi
30
2.8T
104B active at inference time
1M
$2.3
35
KimiBasetenDatabricks
DeepSeek V4 Pro 0424 (Reasoning, High Effort)
DeepSeek logoDeepSeek
30
1.6T
49B active at inference time
1M
$0.2
87
DeepInfraMicrosoft AzureDeepSeek
+5
MiniMax-M3
MiniMax logoMiniMax
30
428B
23B active at inference time
1M
$0.2
130
ParasailCoreWeaveTogether AI
+12