GPT-5.6 Sol vs Gemma 4 E4B: Release Comparison

Comparison of the GPT-5.6 Sol and Gemma 4 E4B releases: 6 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 8 models.

  • For intelligence, GPT-5.6 Sol scores highest: GPT-5.6 Sol (max) at 47, against Gemma 4 E4B (Reasoning) at 9 for Gemma 4 E4B.
  • For output speed, GPT-5.6 Sol is fastest: GPT-5.6 Sol (max) at 65 t/s, against Gemma 4 E4B (Reasoning) at 47 t/s for Gemma 4 E4B.

Releases compared

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench 4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
GPT-5.6 Sol
Pareto line

Cost

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Speed & Latency

Output Speed

Output tokens per second · Higher is better

Capability Scores

Capability Indices

Measures the performance of models on specific capabilities and industries
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench 4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better

Further details

Input modality
Output modality
Weights
Provider Benchmarks
GPT-5.6 Sol (max)
OpenAI logoOpenAI
47
49
55
50
46
49
53
$1.99
$3.1
$4.00
$20.00
$0.40
$3,465
29k
17k
90M
65
94.96s
94.96s
102.69s
450.44s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
Amazon BedrockOpenAIMicrosoft Azure
GPT-5.6 Sol (xhigh)
OpenAI logoOpenAI
44
47
52
49
43
45
52
$1.18
$3.1
$4.00
$20.00
$0.40
$2,082
20k
10k
51M
64
36.15s
36.15s
43.93s
308.51s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
OpenAIMicrosoft Azure
GPT-5.6 Sol (high)
OpenAI logoOpenAI
42
46
51
48
42
43
50
$0.81
$3.1
$4.00
$20.00
$0.40
$1,487
13k
6k
34M
57
9.67s
9.67s
18.38s
229.62s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
OpenAIMicrosoft Azure
GPT-5.6 Sol (medium)
OpenAI logoOpenAI
39
43
49
45
39
40
48
$0.50
$3.1
$4.00
$20.00
$0.40
$997
8k
3k
21M
60
3.39s
3.39s
11.77s
131.41s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
OpenAI
GPT-5.6 Sol (low)
OpenAI logoOpenAI
34
39
42
42
36
35
45
$0.26
$3.1
$4.00
$20.00
$0.40
$637
4k
837
13M
59
3.13s
3.13s
11.54s
64.97s
-
1M
Jul 2026
-
Yes

Supports: text and image

Supports: text

Not available
-
OpenAI
GPT-5.6 Sol (Non-reasoning)
OpenAI logoOpenAI
28
30
33
31
-
-
32
-
$3.1
$4.00
$20.00
$0.40
-
-
-
-
63
1.23s
1.23s
9.16s
-
-
1M
Jul 2026
-
No

Supports: text and image

Supports: text

Not available
-
OpenAIAmazon Bedrock
Gemma 4 E4B (Reasoning)
Google logoGoogle
9
-
-
-
-
-
-
-
$0.0
$0.02
$0.10
-
-
-
-
-
47
0.82s
43.52s
54.19s
-
8B
4.5B active at inference time
128k
Apr 2026
-
Yes

Supports: text, image, speech, and video

Supports: text

DeepInfra
Gemma 4 E4B (Non-reasoning)
Google logoGoogle
7
-
-
-
-
-
-
-
$0.0
$0.02
$0.10
-
-
-
-
-
44
0.77s
0.77s
12.03s
-
8B
4.5B active at inference time
128k
Apr 2026
-
No

Supports: text, image, speech, and video

Supports: text

DeepInfra