Currently listed through these providers:
Model details
GPT-5 Nano
GPT-5 Nano sits at the compact end of the GPT-5 line, designed for situations where speed and responsiveness matter more than deep multi-step reasoning. It carries the core instruction-following behavior of the broader family while shedding the heavier reasoning depth found in larger siblings, making it well suited to rapid-fire chat assistants, code completion helpers, classification passes, and other short-turn developer tasks. As the named successor to GPT-4.1-nano, it inherits that lightweight positioning and extends it into the GPT-5 generation's tooling and safety expectations.
The model's intended sweet spot is real-time, high-volume integrations where every millisecond and every fraction of a cent per token counts. OpenRouter's catalog describes it as optimized for developer tools and ultra-low-latency environments, framing it as a practical choice for product teams that want GPT-5-era capabilities without paying flagship-tier inference costs. Because it accepts text and image inputs while emitting text, teams can route mixed-modality requests through it for triage, summarization, or routing decisions before escalating harder prompts to a larger model in the same family.
Quick Info
Powered by- Provider
- OpenCode Zen
- Model key
- gpt-5-nano
- Release date
- Aug 7, 2025
- Last updated
- Aug 7, 2025
- Knowledge cutoff
- 2024-05-30
- AI SDK package
@ai-sdk/openai- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.05
- Output token cost
- $0.40
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Transparent token rates
Compare gpt-nano pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GPT-5 Nano
Videos about GPT-5 Nano
More models around GPT-5 Nano
This exact model name is also listed by 27 other providers.