Google: Models Intelligence, Performance & Price

Google
Google

This analysis is intended to support you in choosing the best model provided by Google for your use-case.

Most Intelligent

#1
Claude Opus 5.5 (max with fallback)Claude Opus 5.5 (max with fallback)
58
#2
Claude Sonnet 5.5 (max with fallback)Claude Sonnet 5.5 (max with fallback)
56
#3
Claude Opus 5 (max)Claude Opus 5 (max)
51
#4
Claude Opus 5 (xhigh)Claude Opus 5 (xhigh)
50
#5
Claude Fable 5 (with fallback)Claude Fable 5 (with fallback)
50

Intelligence index

Total 63 models

Fastest

#1
Gemini 3.5 Flash-Lite AI StudioGemini 3.5 Flash-Lite AI Studio
359 t/s
#2
gpt-oss-20b (high) Vertexgpt-oss-20b (high) Vertex
306 t/s
#3
Gemini 3.7 Flash (low) AI StudioGemini 3.7 Flash (low) AI Studio
289 t/s
#4
gpt-oss-20b (low) Vertexgpt-oss-20b (low) Vertex
279 t/s
#5
Gemini 2.5 Flash-Lite (AI Studio)Gemini 2.5 Flash-Lite (AI Studio)
277 t/s

Output speed

Total 63 models

Lowest Price

#1
Gemini 2.5 Flash-Lite (AI Studio)Gemini 2.5 Flash-Lite (AI Studio)
$0.07
#2
Gemini 2.5 Flash-Lite (non-reasoning) (AI Studio)Gemini 2.5 Flash-Lite (non-reasoning) (AI Studio)
$0.07
#3
gpt-oss-20b (low) Vertexgpt-oss-20b (low) Vertex
$0.09
#4
gpt-oss-20b (high) Vertexgpt-oss-20b (high) Vertex
$0.09
#5
gpt-oss-120b (high) Vertexgpt-oss-120b (high) Vertex
$0.12

Blended price (per 1M tokens)

Total 63 models

Google offers 63 models, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across models.

  • For intelligence, the top models on Google are Claude Opus 5.5 (max with fallback) (58), Claude Sonnet 5.5 (max with fallback) (56), and Claude Opus 5 (max) (51).
  • For output speed, the fastest models are Gemini 3.5 Flash-Lite AI Studio (359 t/s), gpt-oss-20b (high) Vertex (306 t/s), and Gemini 3.7 Flash (low) AI Studio (289 t/s).
  • For latency, Gemini 2.5 Flash-Lite (non-reasoning) (AI Studio) (0.32s), Gemini 2.5 Flash (non-reasoning) (AI Studio) (0.45s), and Gemini 2.5 Flash (non-reasoning) (Vertex) (0.61s) offer the lowest time to first answer token.
  • For pricing, Gemini 2.5 Flash-Lite (AI Studio) ($0.07), Gemini 2.5 Flash-Lite (non-reasoning) (AI Studio) ($0.07), and gpt-oss-20b (low) Vertex ($0.09) offer the lowest blended prices per 1M tokens.
  • For context window size, Llama 4 Scout Vertex (1M), Claude Opus 5.5 (max with fallback) (1M), and Claude Sonnet 5.5 (max with fallback) (1M) support the largest context windows on Google.
Artificial Analysis Intelligence Index · Higher is better
Output tokens per second · Higher is better
USD per 1M tokens (blended) · Lower is better

Intelligence Evaluations

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3.2 incorporates 10 evaluations: AA-Briefcase v1.1, GDPval-AA v2.1, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Evaluations

Intelligence evaluations measured independently by Artificial Analysis · Higher is better
See more

Agentic knowledge work, (Elo-500)/2000

Agentic real-world work tasks, (Elo-500)/2000

Agentic SaaS workflows

Agentic coding & terminal use

SciCodeUnder review

Coding

Reasoning & knowledge

Professional document reasoning, All-pass

CritPtUnder review

Physics reasoning

Long context reasoning

Legal agentic work, Hallucination-Gated All-Pass Rate

Agentic business operations

Agentic scientific research workflows in a terminal

Quantitative analysis on spreadsheets & documents

Kubernetes incident root-cause analysis

Visual reasoning

Medical long context reasoning

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Context Window

Context Window

Context window: tokens limit · Higher is better

Pricing

Intelligence Index vs. Price

Blended at 7:2:1 (cache-input-output) · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Performance Summary

Output Speed vs. Price

Output speed: output tokens per second · USD per 1M tokens (blended)
Most attractive quadrant
Pareto line

Speed

Measured by Output Speed (tokens per second)

Output Speed

Output tokens per second · Higher is better

Latency

Measured by Time (seconds) to First Token

Latency: Time To First Answer Token

Seconds to first answer token received · Accounts for reasoning model 'thinking' time

Cache Behaviour

Cache Hit Rate

Share of cacheable input tokens served from cache · Median of the last four weeks, updated Sep 26, 2026

Cost per Task vs. Cache Hit Rate

Weighted average cost (USD) per Intelligence Index task · Cache hit rate: median of the last four weeks, updated Sep 26, 2026
Most attractive quadrant
Pareto line

End-to-End Response Time

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

End-to-End Response Time vs. Price

End-to-end response time: end-to-end seconds to output 500 tokens · USD per 1M tokens (blended)
Most attractive quadrant

Further Analysis
Anthropic logo
Claude Opus 5.5 (max with fallback)
1M
Proprietary
58
$15.80
95
685.79
691.04
--
Anthropic logo
Claude Sonnet 5.5 (max with fallback)
1M
Proprietary
56
--
137
499.65
503.30
--
Google logo
Gemini 4 Argon (high) AI Studio (1M output)
1M
Proprietary
53
$1.99
--
--
--
--
Anthropic logo
Claude Opus 5 (max)
1M
Proprietary
51
$9.38
54
56.61
65.82
--
Anthropic logo
Claude Opus 5 (xhigh)
1M
Proprietary
50
$7.75
54
28.19
37.53
--
Anthropic logo
Claude Fable 5 (with fallback)
1M
Proprietary
50
$9.66
60
113.61
121.88
--
Anthropic logo
Claude Opus 5 (high)
1M
Proprietary
48
$5.64
54
18.24
27.46
--
Anthropic logo
Claude Opus 5 (medium)
1M
Proprietary
45
$3.32
54
4.62
13.91
--
Anthropic logo
Claude Opus 5.5 (low with fallback)
1M
Proprietary
42
$0.84
71
6.77
13.77
--
Anthropic logo
Claude Opus 4.8 (max)
1M
Proprietary
42
$7.35
58
27.40
36.06
--
Google logo
Gemini 3.8 Flash (high) AI Studio
1M
Proprietary
41
$1.24
117
25.73
30.02
--
Anthropic logo
Claude Opus 4.7 (max)
1M
Proprietary
41*
--
64
21.08
28.95
--
Google logo
Gemini 3.8 Flash (medium) AI Studio
1M
Proprietary
40
$0.93
--
--
--
--
Google logo
Gemini 3.7 Flash (medium) AI Studio
1M
Proprietary
40*
--
271
6.51
8.35
--
Anthropic logo
Claude Opus 5 (low)
1M
Proprietary
39
$1.57
54
2.15
11.41
--
Google logo
Gemini 3.7 Flash (high) AI Studio
1M
Proprietary
39
$0.93
275
11.35
13.17
--
Anthropic logo
Claude Sonnet 5 (max)
1M
Proprietary
38
$7.36
80
186.48
192.73
--
Google logo
Gemini 3.7 Flash (low) AI Studio
1M
Proprietary
37*
--
283
1.35
3.12
--
Anthropic logo
Claude Sonnet 5.5 (low with fallback)
1M
Proprietary
36
--
99
1.34
6.37
--
Google logo
Gemini 3.6 Flash (high) AI Studio
1M
Proprietary
34
$1.60
189
14.11
16.75
--
Google logo
Gemini 3.5 Flash (medium)
1M
Proprietary
34*
--
195
12.68
15.24
--
Google logo
Gemini 3.8 Flash (low) AI Studio
1M
Proprietary
33
--
--
--
--
--
Google logo
Gemini 3.5 Flash (high) AI Studio
1M
Proprietary
33
$3.69
196
16.60
19.15
--
Anthropic logo
Claude Opus 4.6 (max)
1M
Proprietary
32*
--
42
15.96
27.90
--
Anthropic logo
Claude Sonnet 4.6 (max)
200k
Proprietary
30
$6.47
55
116.65
125.80
--
Google logo
Gemini 3.1 Pro Preview (AI Studio)
1M
Proprietary
30
$1.30
115
23.96
28.29
--
Google logo
Gemini 3.1 Pro Preview (Vertex)
1M
Proprietary
30
$1.31
98
30.04
35.15
--
Anthropic logo
Claude Opus 4.5 Vertex
200k
Proprietary
29*
--
45
14.33
25.39
--
Google logo
Gemini 3 Pro Preview (high) (Vertex)
1M
Proprietary
28*
--
--
--
--
--
Z AI logo
GLM-5
200k
Open
28*
--
69
1.77
53.74
44.76
Anthropic logo
Claude Opus 4.6 (non-reasoning, high)
1M
Proprietary
26*
--
41
1.90
14.14
--
Google logo
Gemini 3 Flash (AI Studio)
1M
Proprietary
26*
--
194
5.83
8.41
--
Anthropic logo
Claude Sonnet 4.6 (non-reasoning, high)
200k
Proprietary
25*
--
42
1.05
12.95
--
Google logo
Gemini 3.5 Flash (minimal) AI Studio
1M
Proprietary
24*
--
193
0.97
3.57
--
Anthropic logo
Claude Opus 4.5 (non-reasoning) Vertex
200k
Proprietary
24*
--
44
0.97
12.26
--
Anthropic logo
Claude 4.1 Opus Vertex
200k
Proprietary
23*
--
--
--
--
--
Google logo
Gemini 3 Pro Preview (low) (Vertex)
1M
Proprietary
22*
--
--
--
--
--
Z AI logo
GLM-4.7
200k
Open
22*
--
180
0.95
14.82
11.10
Google logo
Gemini 3.5 Flash-Lite AI Studio
1M
Proprietary
22
$0.19
359
8.46
9.85
--
Kimi logo
Kimi K2 Thinking Vertex
262k
Open
22*
--
239
0.80
11.24
8.35
Anthropic logo
Claude 4.5 Sonnet Vertex
1M
Proprietary
21
$0.53
39
13.15
25.83
--
Anthropic logo
Claude 4 Opus Vertex
200k
Proprietary
21*
--
--
--
--
--
Anthropic logo
Claude 4.5 Sonnet (non-reasoning) Vertex
1M
Proprietary
19*
--
39
1.04
13.92
--
MiniMax logo
MiniMax-M2 Vertex
197k
Open
19*
--
140
0.62
18.44
14.25
Anthropic logo
Claude 4.1 Opus (non-reasoning) Vertex
200k
Proprietary
19*
--
--
--
--
--
Google logo
Gemini 3 Flash (non-reasoning) (AI Studio)
1M
Proprietary
18*
--
192
1.03
3.64
--
Z AI logo
GLM-4.7 (non-reasoning)
200k
Open
17*
--
172
0.99
3.89
--
Anthropic logo
Claude 4.5 Haiku Vertex
200k
Proprietary
17
$0.24
95
16.65
21.93
--
Google logo
Gemma 4 26B A4B AI Studio
262k
Open
17*
--
46
0.96
54.73
43.01
Anthropic logo
Claude 4 Opus (non-reasoning) Vertex
200k
Proprietary
17*
--
--
--
--
--
Google logo
Gemini 2.5 Pro Vertex
1M
Proprietary
16
$0.33
102
30.64
35.52
--
Google logo
Gemini 2.5 Pro (AI Studio)
1M
Proprietary
16
$0.33
117
23.21
27.48
--
Google logo
Gemini 3.1 Flash-Lite (AI Studio)
1M
Proprietary
16
$0.06
273
5.69
7.52
--
Anthropic logo
Claude 4.5 Haiku (non-reasoning) Vertex
200k
Proprietary
15*
--
83
0.66
6.66
--
Anthropic logo
Claude 3.7 Sonnet (non-reasoning) Vertex
200k
Proprietary
15*
--
--
--
--
--
Google logo
Gemma 4 31B (AI Studio)
262k
Open
15
$0.00
35
1.16
64.70
49.33
Google logo
Gemini 2.5 Pro (May) (AI Studio)
1M
Proprietary
15*
--
--
--
--
--
DeepSeek logo
DeepSeek R1 0528 Vertex
164k
Open
13*
--
149
1.00
17.73
13.39
Google logo
Gemini 2.5 Flash (AI Studio)
1M
Proprietary
13*
--
204
19.99
22.45
--
Google logo
Gemini 2.5 Flash (Vertex)
1M
Proprietary
13*
--
167
21.62
24.62
--
Alibaba logo
Qwen3 235B 2507 Vertex
262k
Open
12*
--
31
0.92
17.13
--
Alibaba logo
Qwen3 Coder 480B Vertex
262k
Open
12*
--
165
0.84
3.86
--
OpenAI logo
gpt-oss-120b (high) Vertex
131k
Open
12
$0.06
42
1.65
60.96
47.45
Alibaba logo
Qwen3 Next 80B A3B Vertex
262k
Open
11*
--
149
0.67
17.43
13.40
Google logo
Gemini 2.5 Flash-Lite (Sep) (AI Studio)
1M
Proprietary
10*
--
--
--
--
--
OpenAI logo
gpt-oss-120b (low) Vertex
131k
Open
10*
--
197
1.53
14.19
10.13
OpenAI logo
gpt-oss-20b (low) Vertex
131k
Open
10*
--
302
0.35
8.62
6.62
Google logo
Gemini 2.5 Flash (non-reasoning) (Vertex)
1M
Proprietary
10*
--
153
0.62
3.89
--
Google logo
Gemini 2.5 Flash (non-reasoning) (AI Studio)
1M
Proprietary
10*
--
187
0.45
3.13
--
Alibaba logo
Qwen3 Next 80B A3B Vertex
262k
Open
10*
--
264
0.64
2.54
--
Google logo
Gemini 2.5 Flash-Lite (Sep) (non-reasoning) (AI Studio)
1M
Proprietary
9*
--
--
--
--
--
OpenAI logo
gpt-oss-20b (high) Vertex
131k
Open
9
$0.01
317
0.34
8.23
6.31
Google logo
Gemini 2.5 Flash-Lite (AI Studio)
1M
Proprietary
9*
--
277
19.18
20.98
--
Google logo
Gemini 2.0 Flash (exp) (AI Studio)
1M
Proprietary
8*
--
--
--
--
--
Meta logo
Llama 4 Scout Vertex
1.31M
Open
8*
--
154
0.73
3.97
--
Meta logo
Llama 3.3 70B Vertex
128k
Open
8*
--
162
0.67
3.76
--
Google logo
Gemini 2.5 Flash-Lite (non-reasoning) (AI Studio)
1M
Proprietary
7*
--
241
0.32
2.40
--
Google logo
Gemma 3 27B (AI Studio)
128k
Open
5
$0.00
--
--
--
--
Google logo
Gemma 3 4B (AI Studio)
128k
Open
5*
--
--
--
--
--
Google logo
Gemma 3n E2B (AI Studio)
32k
Open
5*
--
--
--
--
--
Google logo
Gemma 3 1B (AI Studio)
32k
Open
5*
--
--
--
--
--
Google logo
Gemma 3 12B (AI Studio)
128k
Open
4
$0.00
--
--
--
--

Key definitions

Frequently Asked Questions

Common questions about Google