GPT-5.6 Terra vs Grok 4.3: Release Comparison

Comparison of the GPT-5.6 Terra and Grok 4.3 releases: 6 and 4 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 10 models.

  • For intelligence, GPT-5.6 Terra scores highest: GPT-5.6 Terra (max) at 42, against Grok 4.3 (high) at 25 for Grok 4.3.
  • For output speed, Grok 4.3 is fastest: Grok 4.3 (high) at 119 t/s, against GPT-5.6 Terra (max) at 85 t/s for GPT-5.6 Terra.
  • For cost per task, GPT-5.6 Terra and Grok 4.3 tie at $0.14 (GPT-5.6 Terra (low) and Grok 4.3 (Non-reasoning)).

Releases compared

Intelligence

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index v4.3 incorporates 10 evaluations: AA-Briefcase, GDPval-AA v2, AutomationBench-AA, Terminal-Bench v4.0, SciCode, Humanity's Last Exam, GDP.pdf, CritPt, AA-Omniscience, AA-LCR v1.1
Estimate (independent evaluation forthcoming)

Intelligence Index vs. Cost per Intelligence Index Task

Artificial Analysis Intelligence Index · Weighted average cost (USD) per Artificial Analysis Intelligence Index task
Most attractive quadrant
GPT-5.6 Terra
Grok 4.3
Pareto line

Cost

Cost per Intelligence Index Task

Weighted average cost (USD) per Artificial Analysis Intelligence Index task, segmented by token type. Lower is better

Speed & Latency

Output Speed

Output tokens per second · Higher is better

Capability Scores

Capability Indices

Measures the performance of models on specific capabilities and industries
Finance & Accounting Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Strategy & Ops Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better

Legal Index

Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better

Healthcare & Medical Index

Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better

Engineering Index

Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench v4.0 · Higher is better

Economics Index

Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better

Further details

Weights
Provider Benchmarks
GPT-5.6 Terra (max)
OpenAI logoOpenAI
42
-
1M
$1.7
85
Not available
OpenAIAmazon Bedrock
GPT-5.6 Terra (xhigh)
OpenAI logoOpenAI
38
-
1M
$1.7
75
Not available
OpenAIMicrosoft Azure
GPT-5.6 Terra (high)
OpenAI logoOpenAI
34
-
1M
$1.7
78
Not available
OpenAIMicrosoft Azure
GPT-5.6 Terra (medium)
OpenAI logoOpenAI
30
-
1M
$1.7
76
Not available
OpenAI
GPT-5.6 Terra (low)
OpenAI logoOpenAI
28
-
1M
$1.7
74
Not available
OpenAI
Grok 4.3 (high)
SpaceXAI logoSpaceXAI
25
-
1M
$0.6
119
Not available
SpaceXAIAmazon BedrockMicrosoft Azure
Grok 4.3 (medium)
SpaceXAI logoSpaceXAI
25
-
1M
$0.6
114
Not available
Amazon BedrockMicrosoft AzureSpaceXAI
Grok 4.3 (low)
SpaceXAI logoSpaceXAI
24
-
1M
$0.6
112
Not available
Amazon BedrockMicrosoft AzureSpaceXAI
GPT-5.6 Terra (Non-reasoning)
OpenAI logoOpenAI
22
-
1M
$1.7
74
Not available
OpenAIAmazon Bedrock
Grok 4.3 (Non-reasoning)
SpaceXAI logoSpaceXAI
15
-
1M
$0.6
110
Not available
Amazon BedrockSpaceXAI