Currently listed through these providers:
Model details
GPT-5.4 mini
GPT-5.4 mini sits in the middle of the GPT-5.4 lineup as a smaller, faster sibling designed for workloads where response time shapes the user experience, such as coding assistants, coding subagents, computer-using systems that interpret screenshots, and multimodal applications reasoning over images in real time. OpenAI framed it as a substantial step up from GPT-5 mini, reporting meaningful gains across coding, reasoning, multimodal understanding, and tool use while running more than twice as fast. That combination positions the model as a practical choice when teams want near-frontier capability without paying for the largest model on every request.
Benchmark evidence published alongside the release suggests GPT-5.4 mini gets close to the full-size GPT-5.4 on several professional evaluations, including SWE-Bench Pro and OSWorld-Verified, while still leaving room for the larger model on the hardest tasks. In the GPT-5.4 family, mini takes the balanced tier for high-volume coding and agentic work, with nano reserved for the cheapest, lowest-latency jobs and the flagship GPT-5.4 reserved for the most demanding calls. The trade-off is intentional: for subagent supporting tasks, classification, ranking, and tool-heavy pipelines, mini aims to be the model that responds quickly, uses tools reliably, and still handles complex professional work.
Quick Info
Powered by- Provider
- Requesty
- Model key
- gpt-5.4-mini
- Release date
- Mar 17, 2026
- Last updated
- Mar 17, 2026
- Knowledge cutoff
- 2025-08-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.75
- Output token cost
- $4.50
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Latest news about GPT-5.4 mini
Videos about GPT-5.4 mini
Recent tweets and retweets from Requesty
More models around GPT-5.4 mini
This exact model name is also listed by 33 other providers.