小型开放权重 AI 模型比较(4B-40B)
参数量介于 4B 和 40B 之间的开放权重 AI 模型。
如果模型权重可供下载,我们便将其视为开放权重模型。这样用户就可以在自己的基础设施上自行托管,并通过微调等方式定制模型。
如需了解包括方法论在内的更多详情,请参阅常见问题。
亮点
开放性
Artificial Analysis Openness Index: Score
Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)
Reasoning models are indicated by a lightbulb icon
智能
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index v4.1.1 incorporates 9 evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR
Reasoning models are indicated by a lightbulb icon
Intelligence Evaluations
Intelligence evaluations measured independently by Artificial Analysis · Higher is better
Agentic real-world work tasks, (Elo-500)/2000
𝜏³-BankingUpdated
Agentic tool use
Agentic coding & terminal use
Coding
Humanity's Last ExamUpdated
Reasoning & knowledge
Scientific reasoning
Physics reasoning
AA-Omniscience AccuracyUpdated
Knowledge
1 - hallucination rate
AA-LCRUpdated
Long context reasoning
Agentic knowledge work, Elo
Agentic SaaS workflows
No data available
Legal agentic work, criterion pass rate
Agentic business operations
Quantitative analysis on spreadsheets & documents
No data available
Instruction following
Long-horizon agentic tasks
No data available
Kubernetes incident root-cause analysis
Visual reasoning
Reasoning models are indicated by a lightbulb icon
规模
Model Size: Total and Active Parameters
Comparison between total model parameters and parameters active during inference
Reasoning models are indicated by a lightbulb icon
Intelligence Index vs. Active Parameters
Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line
Reasoning models are indicated by a lightbulb icon
Intelligence Index vs. Total Parameters
Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line
Reasoning models are indicated by a lightbulb icon
上下文窗口
Context Window
Context window: tokens limit · Higher is better
Reasoning models are indicated by a lightbulb icon
更多详情
权重 | 服务商基准测试 | ||||||||
|---|---|---|---|---|---|---|---|---|---|
Qwen3.8 27B (xhigh) | 52 | 27B | 256k | $0.4 | 54 | +2 | |||
Qwen3.8 27B (medium) | 44 | 27B | 256k | $0.4 | 61 | ||||
Qwen3.8 27B (low) | 43 | 27B | 256k | $0.4 | 64 | ||||
Qwen3.6 27B (Reasoning) | 38 | 27.8B | 262k | $0.9 | 56 | +3 | |||
Muse Glimmer (high) | 35 | 30B | 131k | $0.2 | 109 | ||||
Qwen3.8 27B (Non-reasoning) | 35 | 27B | 256k | $0.4 | 61 | ||||
G9v3-39A5B | 34 | 39B 推理时启用 5B 个参数 | 131k | - | - | ||||
Qwen3.6 35B A3B (Reasoning) | 32 | 36B 推理时启用 3B 个参数 | 262k | $0.6 | 122 | +5 | |||
Qwen3.6 27B (Non-reasoning) | 31 | 27.8B | 262k | $0.9 | 58 | +2 | |||
Gemma 4 31B (Reasoning) | 30 | 30.7B | 256k | - | 35 | +11 | |||
Gemma 4 26B A4B (Reasoning) | 26 | 25.2B 推理时启用 3.8B 个参数 | 256k | $0.1 | - | +5 | |||
Qwen3.6 35B A3B (Non-reasoning) | 25 | 36B 推理时启用 3B 个参数 | 262k | $0.6 | 140 | +4 |