This model is deprecated. We only continue performance benchmarking for the default 10k input token workload. Results for other workloads are historical and no longer updated.
SpaceXAI has launched a newer model, Grok 4. We suggest considering it instead.
For more information, see comparison of Grok 4 to other models and API provider benchmarks for Grok 4.
Grok 3 Intelligence, Performance & Price Analysis
Model summary
Intelligence
Speed
Price
Cache Hit Price
Verbosity
Grok 3 is below average in intelligence and particularly expensive when comparing to other non-reasoning models of similar price. The model supports text input, outputs text and image, and has a 1M tokens context window.
Grok 3 scores 18 on the Artificial Analysis Intelligence Index, placing it below average among comparable models (median: 19).
Pricing for Grok 3 is $4.00 per 1M input tokens (expensive, median: $2.00) and $20.00 per 1M output tokens (expensive, median: $8.00).
| Reasoning | No This page shows the non-reasoning version of this model. A reasoning variant may also exist. |
|---|---|
| Input modality | Supports: text |
| Output modality | Supports: text and image |
| Context window | 1M ~1500 A4 pages of size 12 Arial font |
Metrics are compared against models of the same class:
- Non-reasoning models → compared only with other non-reasoning models
- Reasoning models → compared across both reasoning and non-reasoning
- Open weights models → compared only with other open weights models of the same size class:
- Tiny: ≤4B parameters
- Small: 4B–40B parameters
- Medium: 40B–150B parameters
- Large: >150B parameters
- Proprietary models → compared across proprietary and open weights models of the same price range, using a blended 3:1 input/output price ratio:
- <$0.15 per 1M tokens
- $0.15–$1 per 1M tokens
- >$1 per 1M tokens
Highlights
Speed
Intelligence
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index by Open Weights / Proprietary
Intelligence Evaluations
Agentic real-world work tasks, (Elo-500)/2000
Agentic tool use
Agentic coding & terminal use
Coding
Reasoning & knowledge
Scientific reasoning
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Agentic knowledge work, Elo
Agentic SaaS workflows
Legal agentic work, criterion pass rate
Agentic business operations
Instruction following
Long-horizon agentic tasks
Kubernetes incident root-cause analysis
Visual reasoning
AA-Omniscience
AA-Omniscience Index
Intelligence Index Comparisons
Intelligence Index vs. Cost per Intelligence Index Task
Price and Cost
Cost per Intelligence Index Task
Cost to Run Artificial Analysis Intelligence Index
Pricing: Cache Hit, Input, and Output
Context Window
Context Window
Frequently Asked Questions
Common questions about Grok 3
Grok 3 was released on February 19, 2025.
Grok 3 was created by SpaceXAI.
Grok 3 scores 18 (estimated) on the Artificial Analysis Intelligence Index, placing it below average among other non-reasoning models in a similar price tier (median: 19).
Grok 3 costs $4.00 per 1M input tokens (at the higher end, median: $2.00) and $20.00 per 1M output tokens (at the higher end, median: $8.00), based on the median across providers serving the model.
Grok 3 costs $4.00 per 1M input tokens and $20.00 per 1M output tokens (based on the median across providers serving the model). For a blended rate (7:2:1 cache hit/input/output ratio), this is $3.88 per 1M tokens. Pricing may vary by provider. Compare provider pricing
No, Grok 3 is not a reasoning model. It provides direct responses without extended chain-of-thought reasoning.
Grok 3 supports text input.
Grok 3 supports text and image output.
No, Grok 3 does not support image input. It can only process text.
No, Grok 3 is not multimodal. It only supports text input.
Grok 3 has a context window of 1.0M tokens. This determines how much text and conversation history the model can process in a single request.
No, Grok 3 is proprietary. The model weights are not publicly available.
Grok 3 is a proprietary model and SpaceXAI has not disclosed the model size or parameter count.
Grok 3 achieves a score of 18 on the Artificial Analysis Intelligence Index. This composite benchmark evaluates models across reasoning, knowledge, mathematics, and coding.
Yes, Grok 3 is available via API through 2 providers. Compare API providers
Grok 3 is available through 2 API providers. Compare providers