오픈 웨이트 모델 비교

오픈 웨이트 AI 모델을 품질, 성능, 추론 속도, 컨텍스트 창, 파라미터 수, 라이선스 세부 정보 등 주요 성능 지표로 비교하고 분석합니다.

웨이트를 다운로드할 수 있는 모델을 오픈 웨이트 모델(흔히 오픈 소스라고도 함)로 간주합니다. 자체 인프라에서 호스팅할 수 있으며 미세 조정 등을 통해 모델을 맞춤 설정할 수 있습니다.

방법론에 관한 자세한 내용은 FAQ에서 확인하세요.

Z AI 로고GLM-5.3 (max)Kimi 로고Kimi K3 (max)은 오픈 웨이트 모델 중 지능이 가장 높으며 Z AI 로고GLM-5.3-FlashAlibaba 로고Qwen3.8 2.4T A95B이 뒤를 잇습니다.

주요 내용

Artificial Analysis Openness Index · Higher is better
Updated
Artificial Analysis Intelligence Index · Higher is better
학습 가능한 파라미터 수(단위: 십억 개)

개방성

Artificial Analysis Openness Index: Score

Openness Index assesses model openness on a 0 to 100 normalized scale (higher is more open)

오픈 웨이트 모델의 발전

Progress in Open Weights vs. Proprietary Intelligence

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

연구소별 오픈 웨이트 언어 모델 지능 추이

크기별 오픈 웨이트 모델 지능 추이

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

지능

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

Instruction following

Agentic tool use

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

크기

모델 크기별 Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Model Size: Total and Active Parameters

Comparison between total model parameters and parameters active during inference

Intelligence Index vs. Active Parameters

Artificial Analysis Intelligence Index · Active parameters at inference time
Most attractive quadrant
Pareto line

Intelligence Index vs. Total Parameters

Artificial Analysis Intelligence Index · Size in parameters (billions)
Most attractive quadrant
Pareto line

컨텍스트 창

Context Window

Context window: tokens limit · Higher is better

상세 정보

웨이트
제공업체 벤치마크
GLM-5.3 (max)
Z AI 로고Z AI
45
753B
추론 시 40B 활성
1M
$0.9
73
ZaiSelf-hostedDeepInfra
+14
Kimi K3 (max)
Kimi 로고Kimi
44
2.8T
추론 시 104B 활성
1M
$2.3
35
ModalTogether AIDigitalOcean
+14
GLM-5.3-Flash
Z AI 로고Z AI
42
320B
추론 시 18B 활성
1M
$0.1
114
ZaiBasetenNovita
+16
Qwen3.8 2.4T A95B
Alibaba 로고Alibaba
40
2.4T
추론 시 95B 활성
984k
$1.2
41
FireworksTogether AIDeepInfra
+3
DeepSeek V4.1 Flash (Reasoning, Max Effort)
DeepSeek 로고DeepSeek
40
552B
추론 시 16B 활성
1M
$0.2
222
DeepSeekFireworksParasail
+2
DeepSeek V4 Pro 0813 (Reasoning, Max Effort)
DeepSeek 로고DeepSeek
36
1.6T
추론 시 49B 활성
1M
$0.7
91
DeepSeekGMISiliconFlow
+7
Qwen3.8 27B (xhigh)
Alibaba 로고Alibaba
34
27B
256k
$0.4
46
DeepInfraSelf-hostedCoreWeave
+5
K2 Horizon 375B A23B
MBZUAI Institute of Foundation Models 로고MBZUAI Institute of Foundation Models
31
375B
추론 시 23B 활성
524k
-
-
-
MiniMax-M3
MiniMax 로고MiniMax
30
428B
추론 시 23B 활성
1M
$0.2
103
ParasailCoreWeaveTogether AI
+12
Inkling (xhigh)
Thinking Machines 로고Thinking Machines
26
975B
추론 시 41B 활성
1M
$0.7
88
Self-hostedDeepInfraBaseten
+3
Nemotron 3 Ultra 550B A55B (Reasoning)
NVIDIA 로고NVIDIA
23
550B
추론 시 55B 활성
262k
$0.5
204
CoreWeaveGMIDeepInfra
+6
Muse Glimmer (high)
Meta 로고Meta
18
30B
131k
$0.2
91
Together AIDeepInfraFireworks
Mistral Medium 3.5
Mistral 로고Mistral
15
128B
256k
$1.2
147
MistralSelf-hosted
gpt-oss-120b (high)
OpenAI 로고OpenAI
12
117B
추론 시 5.1B 활성
131k
$0.2
226
GroqCloudflareScaleway
+17