Grok 4.6 vs Qwen3.5 397B A17B: Release Comparison
Comparison of the Grok 4.6 and Qwen3.5 397B A17B releases: 4 and 2 models respectively, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across all 6 models.
- For intelligence, Grok 4.6 scores highest: Grok 4.6 (high) at 44, against Qwen3.5 397B A17B (Non-reasoning) at 21 for Qwen3.5 397B A17B.
- For output speed, Qwen3.5 397B A17B is fastest: Qwen3.5 397B A17B (Reasoning) at 86 t/s, against Grok 4.6 (medium) at 58 t/s for Grok 4.6.
- For cost per task, Qwen3.5 397B A17B is cheapest: Qwen3.5 397B A17B (Reasoning) at $0.47, against Grok 4.6 (low) at $0.48 for Grok 4.6.
Releases compared
Intelligence
Artificial Analysis Intelligence Index
Intelligence Index vs. Cost per Intelligence Index Task
Cost per Task (USD, Log Scale)
Cost
Cost per Intelligence Index Task
Speed & Latency
Output Speed
Capability Scores
Capability Indexes
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, AutomationBench-AA, AA-LCR v1.1, GDP.pdf · Higher is better
Incorporates 7 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, Humanity's Last Exam, AA-LCR v1.1, GDP.pdf, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, GDPval-AA v2.1, AA-Briefcase v1.1, MLCR-AA, Humanity's Last Exam, AutomationBench-AA · Higher is better
Incorporates 6 evaluations: AA-Omniscience, Humanity's Last Exam, CritPt, GDPval-AA v2.1, AA-Briefcase v1.1, Terminal-Bench 4.0 · Higher is better
Incorporates 5 evaluations: AA-Omniscience, Humanity's Last Exam, GDPval-AA v2.1, AA-Briefcase v1.1, AA-LCR v1.1 · Higher is better
Further details
Input modality | Output modality | Weights | Provider Benchmarks | ||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Grok 4.6 (high) | 44 | 49 | 52 | 52 | 45 | 45 | 54 | $1.86 | $1.4 | $2.00 | $6.00 | $0.50 | $2,352 | 36k | 19k | 94M | 58 | 42.75s | 42.75s | 51.43s | 637.81s | - | 500k | Aug 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Grok 4.6 (xhigh) | 44 | 49 | 51 | 52 | - | 45 | 55 | $2.32 | $1.4 | $2.00 | $6.00 | $0.50 | $2,830 | 38k | 20k | 97M | 55 | 52.42s | 52.42s | 61.46s | 690.38s | - | 500k | Aug 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Grok 4.6 (medium) | 43 | 47 | 49 | 49 | - | 43 | 53 | $1.50 | $1.4 | $2.00 | $6.00 | $0.50 | $1,937 | 28k | 14k | 73M | 58 | 31.86s | 31.86s | 40.43s | 511.17s | - | 500k | Aug 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Grok 4.6 (low) | 35 | 38 | 44 | 44 | - | 35 | 45 | $0.48 | $1.4 | $2.00 | $6.00 | $0.50 | $761 | 10k | 2k | 21M | 55 | 7.98s | 7.98s | 17.07s | 188.35s | - | 500k | Aug 2026 | - | Yes | Supports: text and image | Supports: text | Not available | - | |||
| Qwen3.5 397B A17B (Non-reasoning) | 21 | - | - | - | - | - | - | - | $0.9 | $0.60 | $3.60 | - | - | - | - | - | 85 | 2.18s | 2.18s | 8.10s | - | 397B 17B active at inference time | 262k | Feb 2026 | - | No | Supports: text and image | Supports: text | +3 | ||||
| Qwen3.5 397B A17B (Reasoning) | 18 | 19 | 15 | 21 | - | 21 | 27 | $0.47 | $0.9 | $0.60 | $3.60 | - | $941 | 19k | 9k | 95M | 86 | 2.26s | 39.48s | 45.32s | 181.37s | 397B 17B active at inference time | 262k | Feb 2026 | - | Yes | Supports: text and image | Supports: text | +5 | ||||