Comparisons of Small Open Source AI Models (4B-40B)

Open source AI models with between 4B to 40B parameters.

Models are considered open source (also commonly referred to as open weights) where their weights are accessible to download. This allows self-hosting on your own infrastructure and enables customizing the model such as through fine-tuning.

For more details including relating to our methodology, see our FAQs.

Alibaba logoQwen3.8 27B (xhigh) and Alibaba logoQwen3.8 27B (medium) are the highest intelligence Small open source models, defined as those with 4B-40B parameters, followed by Alibaba logoQwen3.8 27B (low) & MBZUAI Institute of Foundation Models logoK2 Horizon MoVA 36B A4B.

Highlights

Artificial Analysis Openness Index · Higher is better
Updated
Artificial Analysis Intelligence Index · Higher is better
Trainable parameters in billions

Openness

Artificial Analysis Openness Index: Score

Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

No data available

Agentic tool use

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Size

Model Size: Total and Active Parameters

Comparison between total model parameters and parameters active during inference

Intelligence Index vs. Active Parameters

Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line

Intelligence Index vs. Total Parameters

Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Further details

Weights
Provider Benchmarks
Qwen3.8 27B (xhigh)
Alibaba logoAlibaba
34
27B
256k
$0.4
44
DeepInfraSelf-hostedCoreWeave
+5
Qwen3.8 27B (medium)
Alibaba logoAlibaba
28
27B
256k
$0.4
50
Alibaba CloudMultiverse Computing
Qwen3.8 27B (low)
Alibaba logoAlibaba
26
27B
256k
$0.4
51
Alibaba CloudMultiverse Computing
K2 Horizon MoVA 36B A4B
MBZUAI Institute of Foundation Models logoMBZUAI Institute of Foundation Models
26
36B
4B active at inference time
524k
-
-
-
Qwen3.8 27B (Non-reasoning)
Alibaba logoAlibaba
22
27B
256k
$0.4
50
Alibaba CloudMultiverse Computing
G9v3-39A5B
AI9Stars logoAI9Stars
22
39B
5B active at inference time
131k
-
-
AI9Stars
K2 Horizon 7B
MBZUAI Institute of Foundation Models logoMBZUAI Institute of Foundation Models
21
7B
524k
-
-
-
Qwen3.6 35B A3B (Reasoning)
Alibaba logoAlibaba
19
36B
3B active at inference time
262k
$0.6
127
Alibaba CloudSiliconFlowScaleway
+5
Muse Glimmer (high)
Meta logoMeta
18
30B
131k
$0.2
99
Together AIDeepInfraFireworks
Gemma 4 26B A4B (Reasoning)
Google logoGoogle
17
25.2B
3.8B active at inference time
256k
$0.1
-
MakoraNovitaGoogle
+5
Gemma 4 31B (Reasoning)
Google logoGoogle
15
30.7B
256k
-
35
CoreWeaveModularGMI
+12
Qwen3.6 35B A3B (Non-reasoning)
Alibaba logoAlibaba
15
36B
3B active at inference time
262k
$0.6
131
Self-hostedScalewayDeepInfra
+4