小規模オープンウェイトAIモデルの比較(4B~40B)
パラメーター数が4B~40BのオープンウェイトAIモデルです。
ウェイトをダウンロードできるモデルをオープンウェイト(一般にオープンソースとも呼ばれます)とみなします。独自のインフラストラクチャでセルフホストでき、ファインチューニングなどによるモデルのカスタマイズも可能です。
方法論などの詳細は、よくある質問をご覧ください。
ハイライト
オープン性
Artificial Analysis Openness Index: Score
Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)
Reasoning models are indicated by a lightbulb icon
知能
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index v4.1 incorporates 9 evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience, AA-LCR
Estimate (independent evaluation forthcoming)
Reasoning models are indicated by a lightbulb icon
Intelligence Evaluations
Intelligence evaluations measured independently by Artificial Analysis · Higher is better
Agentic real-world work tasks, (Elo-500)/2000
Agentic tool use
Agentic coding & terminal use
Coding
Reasoning & knowledge
Scientific reasoning
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Agentic knowledge work, Elo
Agentic SaaS workflows
No data available
Legal agentic work, criterion pass rate
Agentic business operations
Instruction following
Long-horizon agentic tasks
No data available
Kubernetes incident root-cause analysis
Visual reasoning
Reasoning models are indicated by a lightbulb icon
規模
Model Size: Total and Active Parameters
Comparison between total model parameters and parameters active during inference
Reasoning models are indicated by a lightbulb icon
Intelligence Index vs. Active Parameters
Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line
Reasoning models are indicated by a lightbulb icon
Intelligence Index vs. Total Parameters
Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line
Reasoning models are indicated by a lightbulb icon
コンテキストウィンドウ
Context Window
Context window: tokens limit · Higher is better
Reasoning models are indicated by a lightbulb icon
詳細
ウェイト | プロバイダーのベンチマーク | ||||||||
|---|---|---|---|---|---|---|---|---|---|
Qwen3.6 27B (Reasoning) | 37 | 27.8B | 262k | $0.9 | 59 | +2 | |||
Qwen3.6 35B A3B (Reasoning) | 32 | 36B 推論時に3Bが有効 | 262k | $0.4 | 140 | +5 | |||
G9v3-39A5B | 31 | 39B 推論時に5Bが有効 | 131k | - | - | ||||
Qwen3.6 27B (Non-reasoning) | 30 | 27.8B | 262k | $0.9 | 57 | ||||
Gemma 4 31B (Reasoning) | 29 | 30.7B | 256k | - | 35 | +10 | |||
Gemma 4 26B A4B (Reasoning) | 26 | 25.2B 推論時に3.8Bが有効 | 256k | $0.1 | - | +5 | |||
Qwen3.6 35B A3B (Non-reasoning) | 24 | 36B 推論時に3Bが有効 | 262k | $0.6 | 163 | +4 | |||
Qwen3.5 35B A3B (Non-reasoning) | 24 | 36B 推論時に3Bが有効 | 262k | $0.4 | 170 | ||||
Gemma 4 12B (Reasoning) | 22 | 12B | 256k | $0.1 | 111 | ||||
Gemma 4 31B (Non-reasoning) | 22 | 30.7B | 256k | $0.2 | 68 | +5 | |||
Qwen3.5 9B (Reasoning) | 21 | 9.7B | 262k | $0.1 | 74 | ||||
Apriel-v1.6-15B-Thinker | 21 | 15B | 128k | - | - |