Currently listed through these providers:
Model details
GPT-5.4 mini
GPT-5.4 mini extends the GPT-5.4 family as a streamlined variant that retains the core capabilities of its larger sibling while targeting faster, more economical inference. It is positioned for production environments where latency and cost matter as much as raw capability, making it well suited to chat applications, coding assistants, and agent workflows that operate at scale. The design intent emphasizes reliable instruction following, solid multi-step reasoning, and consistent behavior across diverse tasks, so teams can deploy it for high-volume scenarios without sacrificing general task quality.
As a multimodal text-and-image model, GPT-5.4 mini fits neatly into agent and assistant pipelines that need to read documents or screenshots alongside natural-language instructions while emitting structured text responses. Its efficiency focus makes it a practical successor for organizations phasing out earlier mini-tier models, and its placement on multi-provider routing platforms means deployments can be tuned for price, latency, or tool-calling precision. The combination of a broad context window with throughput-oriented inference makes it especially attractive for retrieval-augmented assistants and coding copilots that must balance depth of context with responsive interaction.
Quick Info
Powered by- Provider
- CrossModel
- Model key
- openai/gpt-5.4-mini
- Release date
- Mar 17, 2026
- Last updated
- Mar 17, 2026
- Knowledge cutoff
- 2025-08-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.75
- Output token cost
- $4.50
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Latest news about GPT-5.4 mini
Videos about GPT-5.4 mini
More models around GPT-5.4 mini
This exact model name is also listed by 33 other providers.