Currently listed through these providers:
Model details
Grok-4-Fast-Non-Reasoning
Grok-4-Fast-Non-Reasoning is positioned as a streamlined variant within xAI's Grok-4 family, designed for developers and applications that need fast, responsive generation without the added latency of internal reasoning traces. The model ships with support for function calling and structured outputs, enabling it to connect directly with external tools and produce responses in organized, machine-readable formats. This makes it particularly well-suited for workflow automation, agentic pipelines, and integration scenarios where reliability and speed take priority over multi-step deliberation.
The model benefits from xAI's infrastructure across multiple regions including us-east-1 and eu-west-1, with built-in support for prompt caching to reduce redundant processing costs. By omitting the reasoning-then-respond pattern found in heavier variants, Grok-4-Fast-Non-Reasoning delivers quicker turnarounds for single-shot tasks, code generation, document synthesis, and conversational applications. The combination of an expansive context window with efficient token pricing makes it a practical choice for teams building high-volume services that still require access to frontier-level capabilities.
Quick Info
Powered by- Provider
- Poe
- Model key
- xai/grok-4-fast-non-reasoning
- Release date
- Sep 16, 2025
- Last updated
- Sep 16, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.20
- Output token cost
- $0.50
Limits
- Output tokens
- 128,000 tokens
- Context window
- 2,000,000 tokens
Transparent token rates
Compare Grok-4-Fast-Non-Reasoning pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Grok-4-Fast-Non-Reasoning
No articles yet. Fetch the latest news to show it here.