Amazon Bedrock: Models Intelligence, Performance & Price

Amazon Bedrock
Amazon Bedrock

This analysis is intended to support you in choosing the best model provided by Amazon Bedrock for your use-case.

Most Intelligent

Updated
#1
Claude Opus 5.5 (max with fallback)Claude Opus 5.5 (max with fallback)
58
#2
Claude Sonnet 5.5 (max with fallback)Claude Sonnet 5.5 (max with fallback)
56
#3
GPT-6 Astra (max)GPT-6 Astra (max)
53
#4
Claude Opus 5 (max)Claude Opus 5 (max)
51
#5
Claude Opus 5 (xhigh)Claude Opus 5 (xhigh)
50

Intelligence index

Total 86 models

Fastest

#1
Ministral 3 3BMinistral 3 3B
422 t/s
#2
Grok 4.3 (high)Grok 4.3 (high)
359 t/s
#3
Ministral 3 8BMinistral 3 8B
309 t/s
#4
Grok 4.3 (medium)Grok 4.3 (medium)
304 t/s
#5
Nova MicroNova Micro
298 t/s

Output speed

Total 86 models

Lowest Price

#1
Nova MicroNova Micro
$0.03
#2
Gemma 3 4BGemma 3 4B
$0.04
#3
Nova LiteNova Lite
$0.05
#4
NVIDIA Nemotron Nano 9B V2 (non-reasoning)NVIDIA Nemotron Nano 9B V2 (non-reasoning)
$0.08
#5
gpt-oss-20b (low)gpt-oss-20b (low)
$0.09

Blended price (per 1M tokens)

Total 86 models

Amazon offers 86 models, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across models.

  • For intelligence, the top models on Amazon are Claude Opus 5.5 (max with fallback) (58), Claude Sonnet 5.5 (max with fallback) (56), and GPT-6 Astra (max) (53).
  • For output speed, the fastest models are Ministral 3 3B (422 t/s), Grok 4.3 (high) (359 t/s), and Ministral 3 8B (309 t/s).
  • For latency, Grok 4.3 (non-reasoning) (0.63s), GPT-5.6 Luna (non-reasoning) (0.67s), and GPT-5.6 Terra (non-reasoning) (0.68s) offer the lowest time to first answer token.
  • For pricing, Nova Micro ($0.03), Gemma 3 4B ($0.04), and Nova Lite ($0.05) offer the lowest blended prices per 1M tokens. Prices vary up to 3.4x across models.
  • For context window size, Claude Opus 5.5 (max with fallback) (1M), Claude Sonnet 5.5 (max with fallback) (1M), and Claude Opus 5 (max) (1M) support the largest context windows on Amazon.
Updated
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Intelligence Evaluations

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

SciCodeUnder review

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, criterion pass rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Pricing

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Performance Summary

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Speed

Measured by Output Speed (tokens per second)

Output Speed

Output tokens per second · Higher is better

Latency

Measured by Time (seconds) to First Token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

Cache Behaviour

Cache Hit Rate

Share of cacheable input tokens served from cache · Median of the last four weeks, updated Sep 26, 2026

Cost per Task vs. Cache Hit Rate

Weighted average cost (USD) per Intelligence Index task · Cache hit rate: median of the last four weeks, updated Sep 26, 2026
Most attractive quadrant
Pareto line

End-to-End Response Time

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Further Analysis
Anthropic logo
Claude Opus 5.5 (max with fallback)
1M
Proprietary
58
$5.98
--
--
--
--
Anthropic logo
Claude Sonnet 5.5 (max with fallback)
1M
Proprietary
56
--
133
404.57
408.33
--
OpenAI logo
GPT-6 Astra (max)
524k
Proprietary
53
$3.26
93
316.17
321.55
--
Anthropic logo
Claude Opus 5 (max)
1M
Proprietary
51
$5.86
54
90.64
99.98
--
Anthropic logo
Claude Opus 5 (xhigh)
1M
Proprietary
50
$4.88
53
52.33
61.77
--
Anthropic logo
Claude Opus 5 (high)
1M
Proprietary
48
$3.61
52
30.63
40.17
--
OpenAI logo
GPT-5.6 Sol (max)
1M
Proprietary
47
$2.83
99
100.58
105.64
--
Anthropic logo
Claude Opus 5 (medium)
1M
Proprietary
45
$2.19
54
4.98
14.23
--
Kimi logo
Kimi K3 (max)
1M
Open
44
$1.87
177
1.46
15.59
11.31
Anthropic logo
Claude Opus 5.5 (low with fallback)
1M
Proprietary
42
$0.55
59
11.03
19.48
--
OpenAI logo
GPT-5.6 Terra (max)
1M
Proprietary
42
$1.95
147
122.35
125.75
--
Anthropic logo
Claude Opus 4.8 (max)
1M
Proprietary
42
$4.21
64
42.41
50.23
--
Anthropic logo
Claude Opus 4.7 (max)
1M
Proprietary
41*
--
70
23.50
30.62
--
Anthropic logo
Claude Opus 5 (low)
1M
Proprietary
39
$1.10
52
2.38
11.93
--
OpenAI logo
GPT-5.4 (xhigh)
272k
Proprietary
39*
--
140
97.58
101.15
--
OpenAI logo
GPT-5.5 (xhigh)
272k
Proprietary
38
$4.77
121
31.40
35.54
--
Anthropic logo
Claude Sonnet 5 (max)
1M
Proprietary
38
$6.21
75
202.93
209.55
--
OpenAI logo
GPT-5.6 Luna (max)
1M
Proprietary
37
$1.02
231
87.12
89.29
--
OpenAI logo
GPT-5.5 (high)
272k
Proprietary
37
$2.65
108
20.45
25.09
--
Anthropic logo
Claude Sonnet 5.5 (low with fallback)
1M
Proprietary
36
--
106
1.49
6.20
--
OpenAI logo
GPT-5.5 (medium)
272k
Proprietary
34
$1.50
115
7.73
12.08
--
Anthropic logo
Claude Opus 4.6 (max)
1M
Proprietary
32*
--
45
14.12
25.15
--
Anthropic logo
Claude Opus 4.7 (non-reasoning, high)
1M
Proprietary
31*
--
63
1.21
9.12
--
OpenAI logo
GPT-5.5 (low)
272k
Proprietary
31*
--
105
1.33
6.10
--
Anthropic logo
Claude Sonnet 4.6 (max)
200k
Proprietary
30
$2.83
60
73.84
82.18
--
Anthropic logo
Claude Opus 4.5
200k
Proprietary
29*
--
52
15.45
24.98
--
OpenAI logo
GPT-5.6 Sol (non-reasoning)
1M
Proprietary
28*
--
99
0.75
5.79
--
OpenAI logo
GPT-5.4 (low)
272k
Proprietary
28*
--
119
2.74
6.93
--
Anthropic logo
Claude Opus 4.6 (non-reasoning, high)
1M
Proprietary
26*
--
49
1.64
11.90
--
SpaceXAI logo
Grok 4.3 (high)
524k
Proprietary
25
$0.39
343
9.78
11.24
--
SpaceXAI logo
Grok 4.3 (medium)
524k
Proprietary
25*
--
304
6.20
7.85
--
Anthropic logo
Claude Sonnet 4.6 (non-reasoning, high)
200k
Proprietary
25*
--
66
1.32
8.94
--
SpaceXAI logo
Grok 4.3 (low)
524k
Proprietary
24*
--
275
2.81
4.63
--
Anthropic logo
Claude Opus 4.5 (non-reasoning)
200k
Proprietary
24*
--
49
1.55
11.69
--
Kimi logo
Kimi K2.5
256k
Open
23*
--
98
1.29
36.78
30.37
Anthropic logo
Claude Sonnet 5 (non-reasoning)
1M
Proprietary
23*
--
63
1.19
9.08
--
OpenAI logo
GPT-5.5 (non-reasoning)
272k
Proprietary
23*
--
113
0.76
5.18
--
Anthropic logo
Claude 4.1 Opus
200k
Proprietary
23*
--
--
--
--
--
Z AI logo
GLM-4.7
200k
Open
22*
--
59
1.40
43.76
33.88
Kimi logo
Kimi K2 Thinking
256k
Open
22*
--
141
1.22
18.89
14.14
DeepSeek logo
DeepSeek V3.2
128k
Open
21*
--
67
1.41
38.73
29.86
OpenAI logo
GPT-5.6 Terra (non-reasoning)
1M
Proprietary
21
$0.19
116
0.69
4.98
--
Anthropic logo
Claude 4.5 Sonnet
1M
Proprietary
21
$0.56
50
10.45
20.40
--
Anthropic logo
Claude 4.5 Sonnet (non-reasoning)
1M
Proprietary
19*
--
48
1.45
11.90
--
MiniMax logo
MiniMax-M2
205k
Open
19*
--
125
1.13
21.11
15.98
Anthropic logo
Claude 4.1 Opus (non-reasoning)
200k
Proprietary
19*
--
--
--
--
--
OpenAI logo
GPT-5.4 (non-reasoning)
272k
Proprietary
18*
--
112
0.77
5.24
--
Z AI logo
GLM-4.7 (non-reasoning)
200k
Open
17*
--
69
1.45
8.68
--
Anthropic logo
Claude 4.5 Haiku
200k
Proprietary
17
$0.34
126
14.79
18.75
--
DeepSeek logo
DeepSeek V3.2 (non-reasoning)
128k
Open
16*
--
55
1.70
10.73
--
OpenAI logo
GPT-5.6 Luna (non-reasoning)
1M
Proprietary
16
$0.06
197
0.72
3.26
--
Anthropic logo
Claude 4.5 Haiku (non-reasoning)
200k
Proprietary
15*
--
111
0.77
5.28
--
Z AI logo
GLM-4.7-Flash
200k
Open
15*
--
209
1.03
13.02
9.59
Amazon logo
Nova 2.0 Pro Preview (medium)
256k
Proprietary
14*
--
127
14.65
34.27
15.70
SpaceXAI logo
Grok 4.3 (non-reasoning)
524k
Proprietary
14
$0.32
291
0.62
2.34
--
DeepSeek logo
DeepSeek V3.1 (non-reasoning)
128k
Open
14*
--
151
1.61
4.92
--
Amazon logo
Nova 2.0 Omni (medium)
1M
Proprietary
14*
--
--
--
--
--
Amazon logo
Nova 2.0 Omni (medium)
1M
Proprietary
14*
--
--
--
--
--
DeepSeek logo
DeepSeek V3.1
128k
Open
13*
--
152
1.46
17.95
13.19
Amazon logo
Nova 2.0 Lite (high)
1M
Proprietary
13*
--
--
--
--
--
Amazon logo
Nova 2.0 Lite (high)
1M
Proprietary
13*
--
198
16.89
29.54
10.12
Amazon logo
Nova 2.0 Pro Preview (low)
256k
Proprietary
13*
--
123
7.97
28.32
16.28
Amazon logo
Nova 2.0 Lite (medium)
1M
Proprietary
12*
--
184
17.68
31.29
10.89
Alibaba logo
Qwen3 235B 2507
256k
Open
12*
--
78
1.32
7.71
--
Alibaba logo
Qwen3 Coder 480B
262k
Open
12*
--
39
1.65
14.47
--
Amazon logo
Nova 2.0 Lite (low)
1M
Proprietary
12*
--
168
8.31
23.15
11.87
OpenAI logo
gpt-oss-120b (high)
131k
Open
12
$0.11
85
1.04
30.51
23.58
DeepSeek logo
DeepSeek R1 (Jan)
128k
Open
11
$0.22
142
1.10
21.17
16.55
Amazon logo
Nova 2.0 Omni (low)
1M
Proprietary
11*
--
--
--
--
--
Z AI logo
GLM-4.7-Flash (non-reasoning)
200k
Open
11*
--
212
1.03
3.39
--
OpenAI logo
gpt-oss-120b (low)
131k
Open
10*
--
92
1.03
28.11
21.67
Meta logo
Llama 4 Maverick
128k
Open
10*
--
--
--
--
--
Amazon logo
Nova 2.0 Pro Preview (non-reasoning)
256k
Proprietary
10*
--
114
1.00
5.37
--
OpenAI logo
gpt-oss-20b (low)
131k
Open
10*
--
210
7.44
19.34
9.52
Alibaba logo
Qwen3 Coder 30B A3B
262k
Open
10*
--
233
1.06
3.21
--
Mistral logo
Mistral Large 3
256k
Open
9
$0.10
134
1.19
4.91
--
Alibaba logo
Qwen3 Coder Next
128k
Open
9
$0.78
101
1.24
6.19
--
OpenAI logo
gpt-oss-20b (high)
131k
Open
9
$0.01
175
25.26
39.56
11.44
Amazon logo
Nova 2.0 Lite (non-reasoning)
1M
Proprietary
9*
--
179
1.05
3.85
--
Mistral logo
Magistral Small 1.2
128k
Open
9*
--
76
1.57
34.60
26.43
Amazon logo
Nova 2.0 Omni (non-reasoning)
1M
Proprietary
8*
--
--
--
--
--
Meta logo
Llama 4 Scout
128k
Open
8*
--
--
--
--
--
Anthropic logo
Claude 3.5 Sonnet (Oct)
200k
Proprietary
8*
--
--
--
--
--
Meta logo
Llama 3.3 70B
128k
Open
8*
--
--
--
--
--
Alibaba logo
Qwen3 32B (non-reasoning)
32.8k
Open
7*
--
133
1.18
4.95
--
Anthropic logo
Claude 3.5 Sonnet (June)
200k
Proprietary
7*
--
--
--
--
--
Amazon logo
Nova Pro
300k
Proprietary
7*
--
132
1.09
4.87
--
Meta logo
Llama 3.1 8B
128k
Open
7*
--
--
--
--
--
NVIDIA logo
NVIDIA Nemotron Nano 9B V2 (non-reasoning)
131k
Open
7*
--
167
1.17
4.16
--
Mistral logo
Mistral Large 2 (Jul)
128k
Open
7*
--
--
--
--
--
Amazon logo
Nova Lite
300k
Proprietary
7*
--
170
0.98
3.91
--
Meta logo
Llama 3.1 70B Standard
128k
Open
7*
--
--
--
--
--
Meta logo
Llama 3.1 70B Latency Optimized
128k
Open
7*
--
--
--
--
--
Mistral logo
Ministral 3 14B
256k
Open
6
$0.14
228
0.96
3.15
--
AI21 Labs logo
Jamba 1.5 Large
256k
Open
6*
--
--
--
--
--
Amazon logo
Nova Micro
130k
Proprietary
6*
--
298
0.88
2.55
--
NVIDIA logo
NVIDIA Nemotron Nano 12B v2 VL (non-reasoning)
128k
Open
6*
--
201
1.15
3.64
--
Mistral logo
Mistral Large (Feb)
32.8k
Proprietary
6*
--
--
--
--
--
Anthropic logo
Claude 3 Haiku
200k
Proprietary
6*
--
--
--
--
--
Mistral logo
Ministral 3 8B
256k
Open
5
$0.07
317
0.93
2.51
--
Meta logo
Llama 3 70B
8.19k
Open
5*
--
--
--
--
--
AI21 Labs logo
Jamba 1.5 Mini
256k
Open
5*
--
--
--
--
--
Mistral logo
Mixtral 8x7B
32.8k
Open
5*
--
25
1.16
20.99
--
Mistral logo
Mistral 7B
8.19k
Open
5*
--
--
--
--
--
Google logo
Gemma 3 27B
128k
Open
5
$0.47
73
1.41
8.28
--
Mistral logo
Ministral 3 3B
256k
Open
5
$0.04
421
0.94
2.13
--
Google logo
Gemma 3 4B
128k
Open
5*
--
187
1.06
3.74
--
Meta logo
Llama 3 8B
8.19k
Open
5*
--
--
--
--
--
Google logo
Gemma 3 12B
128k
Open
4
$0.24
97
1.32
6.49
--

Key definitions

Frequently Asked Questions

Common questions about Amazon Bedrock