Qwen3.8 27B vs Grok 4.3: Release Comparison
Comparison of the Qwen3.8 27B and Grok 4.3 releases: 4 and 4 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 8 models.
- For intelligence, Qwen3.8 27B scores highest: Qwen3.8 27B (xhigh) at 34, against Grok 4.3 (high) at 25 for Grok 4.3.
- For output speed, Grok 4.3 is fastest: Grok 4.3 (high) at 114 t/s, against Qwen3.8 27B (Non-reasoning) at 50 t/s for Qwen3.8 27B.
- For cost per task, Grok 4.3 is cheapest: Grok 4.3 (Non-reasoning) at $0.14, against Qwen3.8 27B (xhigh) at $0.82 for Qwen3.8 27B.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indices
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2, AA-Briefcase, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2, AA-Briefcase, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2, AA-Briefcase, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Qwen3.8 27B (xhigh) | 34 | 35 | 37 | 35 | 35 | 34 | 43 | $0.82 | $0.4 | $0.50 | $3.00 | $0.05 | $1,170 | 67k | 48k | 198M | 44 | 4.10s | 49.36s | 60.68s | 1126.79s | 27B | 256k | Aug 2026 | - | Yes | Supports: text, image, and video | Supports: text | +2 | ||||
| Qwen3.8 27B (medium) | 28 | 29 | 35 | 28 | - | 25 | 31 | $0.90 | $0.4 | $0.50 | $3.00 | $0.05 | $977 | 52k | 32k | 113M | 49 | 3.84s | 44.70s | 54.91s | 988.74s | 27B | 256k | Aug 2026 | - | Yes | Supports: text, image, and video | Supports: text | |||||
| Qwen3.8 27B (low) | 26 | 27 | 30 | 28 | - | 24 | 32 | $0.83 | $0.4 | $0.50 | $3.00 | $0.05 | $876 | 45k | 26k | 82M | 49 | 4.00s | 44.47s | 54.58s | 879.59s | 27B | 256k | Aug 2026 | - | Yes | Supports: text, image, and video | Supports: text | |||||
| Grok 4.3 (high) | 25 | 28 | 18 | 33 | 27 | 31 | 41 | $0.17 | $0.6 | $1.25 | $2.50 | $0.20 | $332 | 18k | 11k | 87M | 114 | 18.32s | 18.32s | 22.71s | 155.89s | - | 1M | Apr 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Grok 4.3 (medium) | 25 | - | - | - | - | - | - | - | $0.6 | $1.25 | $2.50 | $0.20 | - | - | - | - | 108 | 12.57s | 12.57s | 17.21s | - | - | 1M | Apr 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Grok 4.3 (low) | 24 | - | - | - | - | - | - | - | $0.6 | $1.25 | $2.50 | $0.20 | - | - | - | - | 102 | 4.55s | 4.55s | 9.43s | - | - | 1M | Apr 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Qwen3.8 27B (Non-reasoning) | 22 | - | - | - | - | - | - | - | $0.4 | $0.50 | $3.00 | $0.05 | - | - | - | - | 50 | 3.96s | 3.96s | 13.95s | - | 27B | 256k | Aug 2026 | - | No | Supports: text, image, and video | Supports: text | |||||
| Grok 4.3 (Non-reasoning) | 15 | 16 | 14 | 19 | - | 15 | 21 | $0.14 | $0.6 | $1.25 | $2.50 | $0.20 | $203 | 9k | 0 | 12M | 108 | 0.70s | 0.70s | 5.32s | 86.55s | - | 1M | Apr 2026 | - | No | Supports: text and image | Supports: text | Not available | - | |||