This model is deprecated. We only continue performance benchmarking for the default 10k input token workload. Results for other workloads are historical and no longer updated.
SpaceXAI has launched a newer model, Grok 4.3 (Non-reasoning). We suggest considering it instead.
For more information, see comparison of Grok 4.3 (Non-reasoning) to other models and API provider benchmarks for Grok 4.3 (Non-reasoning).
Grok 4.1 Fast (Non-reasoning) API Provider Benchmarking & Analysis
Analysis of API providers for Grok 4.1 Fast (Non-reasoning) across performance metrics including latency (time to first token), output speed (output tokens per second), price and others. API providers benchmarked include .
Fastest
Output speed
Total 0 providers
Lowest Latency
Time to first token
Total 0 providers
Lowest Price
Blended price (per 1M tokens)
Total 0 providers
No API providers are currently available for Grok 4.1 Fast.
Benchmarks of providers are not available for this model.
Please see the models page for Grok 4.1 Fast (Non-reasoning) for details of the model and its intelligence compared to other models.
Highlights
Pricing
Pricing: Cache Hit, Input, and Output
Pricing: Blended Price
Output Speed vs. Price
Speed
Measured by Output Speed (tokens per second)
Output Speed: Grok 4.1 Fast (Non-reasoning)
Latency vs. Output Speed
Latency
Measured by Time (seconds) to First Token
Time to First Token: Grok 4.1 Fast Providers
End-to-End Response Time
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
End-to-End Response Time: Grok 4.1 Fast Providers
Key Comparison Metrics & API Features
| No results. | |||||||||
Frequently Asked Questions
Common questions about Grok 4.1 Fast (Non-reasoning) providers
Grok 4.1 Fast (Non-reasoning) is not currently available through any API providers we benchmark. Check back later for availability updates.