- Organizations
- xAI
- Grok-4 Fast Non-Reasoning
Grok-4 Fast Non-Reasoning: API Pricing, Context Window & Benchmarks
Grok-4 Fast Non-Reasoning is a language model from xAI, released in August 2025, with multimodal input.
Pushing the Frontier of Cost-Efficient Intelligence. Grok-4 Fast Non-Reasoning uses no thinking tokens for immediate responses, making it faster while maintaining high quality.
Grok-4 Fast Non-Reasoning pricing
Providers
Grok-4 Fast Non-Reasoning starts at $0.200 per million input tokens and $0.500 per million output tokens via xAI.
| Provider | Input $/M | Cached input $/M | Output $/M | Context in / out | TTFT p95 s | Output p5 c/s | Modalities in / out |
|---|---|---|---|---|---|---|---|
| $0.200 | — | $0.500 | 2.0M/30.0K | — | — | / |
Cached input is the discounted price for prompt tokens served from a provider cache. TTFT is time to first token. Output is characters per second; p5 is the sustained floor exceeded by 95% of observed requests.
Grok-4 Fast Non-Reasoning context window
Input and output token limits for Grok-4 Fast Non-Reasoning, plus how it ranks on long-context understanding.
Grok-4 Fast Non-Reasoning API
Available from the model provider
Grok-4 Fast Non-Reasoning has an official provider API. It is not currently routed through the LLM Stats gateway.
Read the official API documentationGrok-4 Fast Non-Reasoning latency
Grok-4 Fast Non-Reasoning time to first token, sustained output throughput, and failed-request rate from live API traffic over the trailing 7 days.
Grok-4 Fast Non-Reasoning examples
Recent arena outputs from Grok-4 Fast Non-Reasoning, picked from the highest-ranked matchups.
Grok-4 Fast Non-Reasoning license
Grok-4 Fast Non-Reasoning is a proprietary model available under its provider's product and API terms.
- License
- Proprietary
- Hosted access
Proprietary license - usage restrictions apply
Grok-4 Fast Non-Reasoning resources
Official sources for Grok-4 Fast Non-Reasoning: api documentation.
FAQ
Common questions about Grok-4 Fast Non-Reasoning.