Currently listed through these providers:
Model details
GPT-5 Nano
GPT-5 Nano is the most compact member of the GPT-5 family, engineered specifically for environments where low latency and high throughput are the primary constraints. Unlike its larger siblings, Nano is architected to facilitate rapid, real-time interactions and lightweight agentic tasks, serving as a fast-response engine for routine classifications, basic summarizations, and high-frequency API calls. The model utilizes a dense transformer architecture optimized for efficient scaling and incorporates variable reasoning effort levels—minimal, low, medium, and high—allowing developers to tune the balance between inference speed and cognitive depth per request. This flexibility, combined with an expanded context window, enables the model to process extensive document sets or lengthy conversation histories despite its smaller parameter footprint, while multi-modal input support lets it handle both text and image data.
As the direct successor to GPT-4.1-nano, GPT-5 Nano slots into a tiered ecosystem alongside GPT-5 Mini, GPT-5 Pro, and the flagship GPT-5, each sized for different performance and cost requirements. It retains the instruction-following precision and safety features characteristic of the broader GPT-5 lineage while functioning as part of a unified routing system that dynamically allocates compute resources. The model is optimized for developer tools, ultra-low latency environments, and cost-sensitive applications where the depth of larger variants is unnecessary. This makes it particularly well-suited for high-volume API calls and real-time use cases where speed matters more than extended reasoning, offering a lightweight entry point to the GPT-5 generation without sacrificing the core capabilities developers expect from OpenAI's ecosystem.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- gpt-5-nano
- Release date
- Aug 7, 2025
- Last updated
- Aug 7, 2025
- Knowledge cutoff
- 2024-05-30
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.05
- Output token cost
- $0.40
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Transparent token rates
Compare gpt-nano pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GPT-5 Nano
No articles yet. Fetch the latest news to show it here.
Videos about GPT-5 Nano
More models around GPT-5 Nano
This exact model name is also listed by 27 other providers.