대형 오픈 웨이트 AI 모델 비교(>150B)
파라미터 수가 150B개를 넘는 오픈 웨이트 AI 모델입니다.
웨이트를 다운로드할 수 있는 모델을 오픈 웨이트 모델(흔히 오픈 소스라고도 함)로 간주합니다. 자체 인프라에서 호스팅할 수 있으며 미세 조정 등을 통해 모델을 맞춤 설정할 수 있습니다.
방법론을 비롯한 자세한 내용은 FAQ에서 확인하세요.
주요 내용
개방성
Artificial Analysis Openness Index: Score
Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)
지능
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)
Intelligence Evaluations
Intelligence evaluations measured independently by Artificial Analysis · Higher is better
Agentic knowledge work, (Elo-500)/2000
Agentic real-world work tasks, (Elo-500)/2000
AutomationBench-AAUpdated
Agentic SaaS workflows
Agentic coding & terminal use
Coding
Reasoning & knowledge
GDP.pdfNew
Professional document reasoning, All-pass
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Legal agentic work, criterion pass rate
Agentic business operations
Quantitative analysis on spreadsheets & documents
Instruction following
Agentic tool use
Long-horizon agentic tasks
Kubernetes incident root-cause analysis
Visual reasoning
크기
Model Size: Total and Active Parameters
Comparison between total model parameters and parameters active during inference
Intelligence Index vs. Active Parameters
Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line
Intelligence Index vs. Total Parameters
Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line
컨텍스트 창
Context Window
Context window: tokens limit · Higher is better
상세 정보
웨이트 | 제공업체 벤치마크 | ||||||||
|---|---|---|---|---|---|---|---|---|---|
GLM-5.3 (max) | 45 | 753B 추론 시 40B 활성 | 1M | $0.9 | 72 | +14 | |||
Kimi K3 (max) | 44 | 2.8T 추론 시 104B 활성 | 1M | $2.3 | 35 | +14 | |||
GLM-5.3-Flash | 42 | 320B 추론 시 18B 활성 | 1M | $0.1 | 107 | +16 | |||
Qwen3.8 2.4T A95B | 40 | 2.4T 추론 시 95B 활성 | 984k | $1.2 | 41 | +3 | |||
Qwen3.8-Flash-Next | 40 | 180B 추론 시 6B 활성 | 256k | $0.1 | 52 | ||||
DeepSeek V4.1 Flash (Reasoning, Max Effort) | 40 | 552B 추론 시 16B 활성 | 1M | $0.2 | 214 | +2 | |||
DeepSeek V4 Pro 0813 (Reasoning, Max Effort) | 36 | 1.6T 추론 시 49B 활성 | 1M | $0.7 | 85 | +7 | |||
DeepSeek V4 Flash 0731 (Reasoning, Max Effort) | 35 | 284B 추론 시 13B 활성 | 1M | $0.2 | 213 | +14 | |||
GLM-5.2 (max) | 34 | 753B 추론 시 40B 활성 | 1M | $0.9 | 75 | +18 | |||
Motif 3 | 34 | 314B 추론 시 13.2B 활성 | 262k | - | - | 제공되지 않음 | - | ||
DeepSeek V4 Pro (Reasoning, Max Effort) | 31 | 1.6T 추론 시 49B 활성 | 1M | $0.2 | 82 | +8 | |||
K2 Horizon 375B A23B | 31 | 375B 추론 시 23B 활성 | 524k | - | - | - |