Microsoft Azure: 모델 지능, 성능 및 가격

Microsoft Azure
Microsoft Azure

이 분석은 사용 사례에 가장 적합한 Microsoft Azure 제공 모델을 선택하는 데 도움을 드립니다.

가장 높은 지능

Updated
#1
Claude Fable 5 (with fallback)Claude Fable 5 (with fallback)
50
#2
GPT-5.6 Sol (max)GPT-5.6 Sol (max)
47
#3
GPT-5.6 Sol (xhigh)GPT-5.6 Sol (xhigh)
44
#4
GPT-5.6 Sol (high)GPT-5.6 Sol (high)
42
#5
Claude Opus 4.7 (max)Claude Opus 4.7 (max)
41

Intelligence Index

총 모델 79개

가장 빠름

#1
Llama 4 Maverick (FP8)Llama 4 Maverick (FP8)
462 t/s
#2
gpt-oss-120b (low)gpt-oss-120b (low)
329 t/s
#3
gpt-oss-120b (high)gpt-oss-120b (high)
308 t/s
#4
GPT-5.4 mini (xhigh)GPT-5.4 mini (xhigh)
294 t/s
#5
o3-minio3-mini
266 t/s

출력 속도

총 모델 79개

가장 낮은 가격

#1
GPT-5 nano (high)GPT-5 nano (high)
$0.05
#2
GPT-5 nano (medium)GPT-5 nano (medium)
$0.05
#3
GPT-4.1 nanoGPT-4.1 nano
$0.08
#4
GPT-4o miniGPT-4o mini
$0.14
#5
Phi-4Phi-4
$0.16

혼합 가격(토큰 100만 개당)

총 모델 79개

Azure에서 지능, 성능, 가격 특성이 서로 다른 모델 79개를 제공합니다. 아래에서 모델별 주요 지표를 비교합니다.

  • Azure에서 지능이 가장 높은 모델은 Claude Fable 5 (with fallback)(50), GPT-5.6 Sol (max)(47) 및 GPT-5.6 Sol (xhigh)(44)입니다.
  • 출력 속도가 가장 빠른 모델은 Llama 4 Maverick (FP8)(462 t/s), gpt-oss-120b (low)(329 t/s) 및 gpt-oss-120b (high)(308 t/s)입니다. 모델별 속도 차이가 크며, 가장 빠른 모델과 가장 느린 모델의 차이는 74%입니다.
  • 첫 답변 토큰까지 걸린 시간이 가장 짧은 모델은 Phi-4 Mini(0.86초), Llama 4 Scout(0.87초) 및 Phi-4 Multimodal(0.87초)입니다.
  • 토큰 100만 개당 혼합 가격이 가장 낮은 모델은 GPT-5 nano (high)($0.05), GPT-5 nano (medium)($0.05) 및 GPT-4.1 nano($0.08)입니다. 모델별 가격은 최대 3.0배 차이 납니다.
  • Azure에서 가장 큰 컨텍스트 창을 지원하는 모델은 GPT-5.4 (xhigh)(1M), Claude Fable 5 (with fallback)(1M) 및 Claude Opus 4.7 (max)(1M)입니다.

주요 내용

Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

지능 평가

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Quantitative analysis on spreadsheets & documents

Instruction following

Agentic tool use

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

컨텍스트 창

Context Window

Context window: tokens limit · Higher is better

가격

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

성능 요약

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

속도

출력 속도(초당 토큰 수)로 측정

Output Speed

Output tokens per second · Higher is better

지연 시간

첫 토큰까지 걸린 시간(초)으로 측정

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

종단 간 응답 시간

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

추가 분석
Anthropic 로고
Claude Fable 5 (with fallback)
1M
독점
50
$43.25
56
72.24
81.13
--
OpenAI 로고
GPT-5.6 Sol (max)
524k
독점
47
--
98
90.87
95.96
--
OpenAI 로고
GPT-5.6 Sol (xhigh)
524k
독점
44
$1.26
97
22.24
27.40
--
OpenAI 로고
GPT-5.6 Sol (high)
524k
독점
42
$0.85
97
8.02
13.16
--
Anthropic 로고
Claude Opus 4.7 (max)
1M
독점
41*
--
48
17.55
27.97
--
OpenAI 로고
GPT-5.4 (xhigh)
1.05M
독점
39*
--
114
131.39
135.79
--
Anthropic 로고
Claude Sonnet 5 (max)
1M
독점
38
$4.75
73
170.99
177.84
--
OpenAI 로고
GPT-5.6 Terra (xhigh)
524k
독점
38
$0.67
109
11.55
16.13
--
OpenAI 로고
GPT-5.5 (high)
524k
독점
37
$1.13
105
17.31
22.09
--
OpenAI 로고
GPT-5.6 Luna (xhigh)
524k
독점
35
$0.09
129
50.60
54.47
--
OpenAI 로고
GPT-5.6 Terra (high)
524k
독점
34
$0.35
108
2.44
7.08
--
OpenAI 로고
GPT-5.6 Luna (high)
524k
독점
32
$0.05
121
11.94
16.07
--
Anthropic 로고
Claude Opus 4.6 (max)
1M
독점
32*
--
40
9.09
21.64
--
DeepSeek 로고
DeepSeek V4 Pro (max)
1M
오픈
31
$9.65
93
1.65
53.86
46.86
Anthropic 로고
Claude Sonnet 4.6 (max)
200k
독점
30
$2.70
50
122.86
132.84
--
OpenAI 로고
GPT-5.2 (xhigh)
400k
독점
30*
--
85
97.46
103.32
--
DeepSeek 로고
DeepSeek V4 Pro (high)
1M
오픈
30*
--
88
1.63
29.83
22.54
Anthropic 로고
Claude Opus 4.5
200k
독점
29*
--
45
9.42
20.54
--
OpenAI 로고
GPT-5.2 Codex (xhigh)
400k
독점
29*
--
96
105.53
110.72
--
Kimi 로고
Kimi K2.6
262k
오픈
27
$1.91
226
1.46
23.40
19.73
OpenAI 로고
GPT-5.2 (medium)
400k
독점
27*
--
83
4.84
10.85
--
Anthropic 로고
Claude Opus 4.6 (Non-reasoning, high)
1M
독점
26*
--
37
1.94
15.61
--
SpaceXAI 로고
Grok 4.20 0309 v2
262k
독점
26*
--
225
9.18
11.41
--
SpaceXAI 로고
Grok 4.3 (high)
200k
독점
25
$0.53
168
15.04
18.01
--
OpenAI 로고
GPT-5 Codex (high)
400k
독점
25*
--
169
6.19
9.15
--
SpaceXAI 로고
Grok 4.3 (medium)
200k
독점
25*
--
172
9.56
12.47
--
OpenAI 로고
GPT-5.1 (high)
272k
독점
25*
--
133
30.57
34.32
--
Anthropic 로고
Claude Sonnet 4.6 (Non-reasoning, high)
200k
독점
25*
--
42
2.20
14.07
--
OpenAI 로고
GPT-5.4 mini (xhigh)
400k
독점
25
$0.48
290
126.01
127.73
--
SpaceXAI 로고
Grok 4.3 (low)
200k
독점
24*
--
160
4.71
7.84
--
OpenAI 로고
GPT-5.1 Codex (high)
400k
독점
24*
--
125
7.62
11.61
--
Anthropic 로고
Claude Opus 4.5 (Non-reasoning)
200k
독점
24*
--
44
1.92
13.40
--
Kimi 로고
Kimi K2.6 (Non-reasoning)
262k
오픈
24*
--
179
1.47
4.26
--
Kimi 로고
Kimi K2.5
262k
오픈
23*
--
158
1.53
23.53
18.82
OpenAI 로고
GPT-5 (high)
400k
독점
23*
--
97
47.60
52.74
--
OpenAI 로고
GPT-5 (medium)
400k
독점
23*
--
109
19.92
24.52
--
Anthropic 로고
Claude 4.1 Opus
200k
독점
23*
--
--
--
--
--
SpaceXAI 로고
Grok 4
256k
독점
22*
--
78
7.55
14.00
--
Kimi 로고
Kimi K2 Thinking
256k
오픈
22*
--
154
1.58
17.85
13.02
OpenAI 로고
o3-pro
200k
독점
22*
--
--
--
--
--
Anthropic 로고
Claude 4.5 Sonnet
200k
독점
21
$1.78
40
9.93
22.58
--
DeepSeek 로고
DeepSeek V4 Pro (Non-reasoning)
1M
오픈
21*
--
84
1.75
7.70
--
OpenAI 로고
GPT-5 (low)
400k
독점
21*
--
92
4.99
10.40
--
OpenAI 로고
GPT-5 mini (medium)
400k
독점
21*
--
144
9.11
12.58
--
OpenAI 로고
GPT-5.1 Codex mini (high)
400k
독점
20*
--
157
12.53
15.73
--
OpenAI 로고
o3
200k
독점
20*
--
99
28.26
33.28
--
OpenAI 로고
GPT-5.4 mini (medium)
400k
독점
20*
--
218
3.70
5.99
--
Kimi 로고
Kimi K2.5 (Non-reasoning)
262k
오픈
19*
--
171
1.54
4.47
--
Anthropic 로고
Claude 4.5 Sonnet (Non-reasoning)
200k
독점
19*
--
39
1.77
14.74
--
Anthropic 로고
Claude 4.1 Opus (Non-reasoning)
200k
독점
19*
--
--
--
--
--
SpaceXAI 로고
Grok 4 Fast
2M
독점
18*
--
--
--
--
--
Anthropic 로고
Claude 4.5 Haiku
200k
독점
18
$0.57
83
28.59
34.60
--
OpenAI 로고
GPT-5 mini (high)
400k
독점
17
$0.05
148
52.56
55.94
--
OpenAI 로고
GPT-5.2 (Non-reasoning)
400k
독점
17*
--
81
1.18
7.39
--
OpenAI 로고
o4-mini (high)
200k
독점
17*
--
153
17.37
20.63
--
DeepSeek 로고
DeepSeek V3.2 (Non-reasoning)
128k
오픈
16*
--
163
1.83
4.91
--
Anthropic 로고
Claude 4.5 Haiku (Non-reasoning)
200k
독점
15*
--
79
1.27
7.64
--
OpenAI 로고
o1
200k
독점
15*
--
--
--
--
--
SpaceXAI 로고
Grok 3 mini Reasoning (high)
32k
독점
15*
--
--
--
--
--
OpenAI 로고
GPT-5.1 (Non-reasoning)
400k
독점
13*
--
116
1.28
5.57
--
OpenAI 로고
GPT-5 nano (high)
400k
독점
13*
--
221
56.72
58.99
--
OpenAI 로고
GPT-4.1
1M
독점
13*
--
135
1.70
5.41
--
OpenAI 로고
GPT-5 nano (medium)
400k
독점
12*
--
208
32.22
34.63
--
OpenAI 로고
o3-mini
200k
독점
12*
--
256
5.85
7.81
--
OpenAI 로고
gpt-oss-120b (high)
131k
오픈
12
$0.11
313
0.77
8.77
6.40
SpaceXAI 로고
Grok 3
16k
독점
12*
--
--
--
--
--
OpenAI 로고
GPT-5 (minimal)
400k
독점
11*
--
94
1.65
6.97
--
OpenAI 로고
o1-preview
128k
독점
11*
--
--
--
--
--
OpenAI 로고
GPT-5.4 mini (Non-reasoning)
400k
독점
11*
--
189
1.07
3.71
--
SpaceXAI 로고
Grok 4 Fast (Non-reasoning)
2M
독점
11*
--
--
--
--
--
OpenAI 로고
o3-mini (high)
200k
독점
11
--
225
19.99
22.21
--
OpenAI 로고
gpt-oss-120b (low)
131k
오픈
10*
--
331
0.82
8.38
6.05
OpenAI 로고
GPT-4.1 mini
1M
독점
10*
--
140
1.24
4.81
--
OpenAI 로고
GPT-5 mini (minimal) East US 2 - Global Standard
400k
독점
10*
--
--
--
--
--
Mistral 로고
Mistral Large 3
256k
오픈
10
$0.10
98
1.43
6.52
--
Meta 로고
Llama 4 Maverick (FP8)
128k
오픈
9
$0.04
447
1.21
2.32
--
Mistral 로고
Mistral Medium 3
128k
독점
9*
--
47
2.02
12.57
--
OpenAI 로고
GPT-4o (Nov)
128k
독점
8*
--
150
2.40
5.73
--
OpenAI 로고
GPT-4.1 nano
1M
독점
8*
--
252
1.59
3.58
--
OpenAI 로고
GPT-4o (Aug)
128k
독점
8*
--
149
1.35
4.70
--
Meta 로고
Llama 3.3 70B
128k
오픈
8*
--
116
2.27
6.59
--
OpenAI 로고
GPT-4o (May)
128k
독점
7*
--
156
1.55
4.76
--
OpenAI 로고
GPT-5 nano (minimal) East US 2 - Global Standard
400k
독점
7*
--
--
--
--
--
OpenAI 로고
GPT-4 Turbo
128k
독점
7*
--
127
1.55
5.48
--
Cohere 로고
Command A
256k
오픈
7*
--
42
3.03
14.84
--
OpenAI 로고
GPT-4o mini
128k
독점
7*
--
92
1.84
7.30
--
Meta 로고
Llama 4 Scout
128k
오픈
6
$0.12
132
0.87
4.67
--
Microsoft 로고
Phi-4 Mini
128k
오픈
6*
--
45
0.86
11.98
--
Microsoft 로고
Phi-4
16.4k
오픈
6*
--
40
2.53
14.91
--
Microsoft 로고
Phi-4 Multimodal
128k
오픈
6*
--
17
0.87
30.20
--

주요 용어 정의

자주 묻는 질문

Microsoft Azure에 관한 일반적인 질문