Currently listed through these providers:
Model details
xAI Grok 4 Fast Non-Reasoning
Grok 4 Fast Non-Reasoning is the stripped-down sibling in xAI's Grok 4 family, designed for high-throughput, low-cost inference rather than deep chain-of-thought reasoning. It is delivered as a proprietary, closed-weights model, so weights and full architecture details are not publicly available, but its intended role is clear from the pricing and benchmark posture: a fast, cheap workhorse that can sit behind large context windows for routine summarization, extraction, classification, and short-form generation workloads where the full Grok 4 would be overkill.
In practice, the model trades peak accuracy for affordability. Independent benchmark snapshots show it landing well below the flagship Grok 4 on math, science, and coding evaluations, scoring roughly a third on AIME 2025, around 62% on GPQA Diamond, and about 46% on LiveCodeBench, compared with the full model's 91.7%, 87.5%, and 79% respectively. The payoff is a blended cost roughly 96% lower than Grok 4, paired with a 2,000,000-token context window that makes it well suited to long-document processing, codebase ingestion, and bulk pipeline jobs where breadth and price matter more than squeezing out the last few accuracy points.
Quick Info
Powered by- Provider
- Helicone
- Model key
- grok-4-fast-non-reasoning
- Release date
- Sep 19, 2025
- Last updated
- Sep 19, 2025
- Knowledge cutoff
- 2025-09
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.20
- Output token cost
- $0.50
Limits
- Output tokens
- 2,000,000 tokens
- Context window
- 2,000,000 tokens
Latest news about xAI Grok 4 Fast Non-Reasoning
No articles yet. Fetch the latest news to show it here.