Currently listed through these providers:
Model details
GPT-5.4 mini
GPT-5.4 mini carries forward the core capabilities of the larger GPT-5.4 line in a smaller, faster design aimed at high-volume workloads. It improves on its predecessor in coding, reasoning, multimodal understanding, and tool use, reportedly running more than twice as fast. Its combination of responsiveness and broad task ability makes it practical for coding assistants, supporting subagents, real-time visual applications, and other systems where latency directly affects the user experience.
The model approaches the larger GPT-5.4 on evaluations such as SWE-Bench Pro and OSWorld-Verified, while retaining an efficiency-focused design. That balance suits production chat systems, coding tools, and agent workflows that need multi-step reasoning, reliable instruction following, and image-informed understanding without moving every task to a larger model.
Quick Info
Powered by- Provider
- FreeModel
- Model key
- gpt-5.4-mini
- Release date
- Mar 17, 2026
- Last updated
- Mar 17, 2026
- Knowledge cutoff
- 2025-08-31
- AI SDK package
@ai-sdk/openai-compatible- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.75
- Output token cost
- $4.50
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Transparent token rates
Compare GPT-5.4 mini pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GPT-5.4 mini
Videos about GPT-5.4 mini
More models around GPT-5.4 mini
This exact model name is also listed by 33 other providers.