Comparison of Open Source Models

Comparison and analysis of open source AI models across key performance metrics including quality, performance, inference speed, context window, parameter count & licensing details.

Models are considered open source (also commonly referred to as open weights) where their weights are accessible to download. This allows self-hosting on your own infrastructure and enables customizing the model such as through fine-tuning.

For more details relating to our methodology, see our FAQs.

Xiaomi logoMiMo-V2.6-Pro and Z AI logoGLM-5.3 (max) are the highest intelligence open source models, followed by Kimi logoKimi K3 (max) & Z AI logoGLM-5.3-Flash.
Artificial Analysis Openness Index · Higher is better
Artificial Analysis Intelligence Index · Higher is better
Trainable parameters in billions

Openness

Artificial Analysis Openness Index: Score

Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)

Open Source Progress

Progress in Open Weights vs. Proprietary Intelligence

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Open Source Language Models Intelligence By Lab Over Time

Open Source Models Intelligence By Size Over Time

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

SciCodeUnder review

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Size

Intelligence Index By Model Size

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Model Size: Total and Active Parameters

Comparison between total model parameters and parameters active during inference (billions)

Intelligence Index vs. Active Parameters

Artificial Analysis Intelligence Index · Active parameters at inference time (billions)
Most attractive quadrant
Pareto line

Intelligence Index vs. Total Parameters

Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Further details

Weights
Provider Benchmarks
MiMo-V2.6-Pro
Xiaomi logoXiaomi
46
1.0T
42B active at inference time
1M
$0.2
40
XiaomiDeepInfraPrimaLabs
+2
GLM-5.3 (Max)
Z AI logoZ AI
45
753B
40B active at inference time
1M
$0.9
79
ModalZaiSelf-hosted
+23
Kimi K3 (Max)
Kimi logoKimi
44
2.8T
104B active at inference time
1M
$2.3
42
DatabricksFireworksFireworks
+18
GLM 5.3 Flash
Z AI logoZ AI
42
320B
18B active at inference time
1M
$0.1
51
ModalZaiBaseten
+21
DeepSeek V4.1 Flash (Max)
DeepSeek logoDeepSeek
39
552B
16B active at inference time
1M
$0.2
216
NebiusSelf-hostedSelf-hosted
+20
Qwen3.8 27B (Xhigh)
Alibaba logoAlibaba
34
27B
256k
$0.5
47
DeepInfraSelf-hostedCoreWeave
+9
K2 Horizon 375B A23B
Institute of Foundation Models logoInstitute of Foundation Models
31
375B
23B active at inference time
524k
-
118
Institute of Foundation Models
MiniMax-M3
MiniMax logoMiniMax
29
428B
23B active at inference time
1M
$0.2
94
ParasailCoreWeaveTogether AI
+12
Inkling (Xhigh)
Thinking Machines logoThinking Machines
25
975B
41B active at inference time
1M
$0.7
151
Self-hostedDeepInfraBaseten
+4
Nemotron 3 Ultra 550B A55B (Reasoning)
NVIDIA logoNVIDIA
23
550B
55B active at inference time
262k
$0.5
155
CoreWeaveGMIDeepInfra
+5
Muse Glimmer (High)
Meta logoMeta
17
30B
131k
$0.2
91
Together AIDeepInfraSystalyzeFireworks