Sulat.com
AI models
Requesty logo

Model details

GPT-5.4 mini

GPT-5.4 mini sits in the middle of the GPT-5.4 lineup as a smaller, faster sibling designed for workloads where response time shapes the user experience, such as coding assistants, coding subagents, computer-using systems that interpret screenshots, and multimodal applications reasoning over images in real time. OpenAI framed it as a substantial step up from GPT-5 mini, reporting meaningful gains across coding, reasoning, multimodal understanding, and tool use while running more than twice as fast. That combination positions the model as a practical choice when teams want near-frontier capability without paying for the largest model on every request.

Benchmark evidence published alongside the release suggests GPT-5.4 mini gets close to the full-size GPT-5.4 on several professional evaluations, including SWE-Bench Pro and OSWorld-Verified, while still leaving room for the larger model on the hardest tasks. In the GPT-5.4 family, mini takes the balanced tier for high-volume coding and agentic work, with nano reserved for the cheapest, lowest-latency jobs and the flagship GPT-5.4 reserved for the most demanding calls. The trade-off is intentional: for subagent supporting tasks, classification, ranking, and tool-heavy pipelines, mini aims to be the model that responds quickly, uses tools reliably, and still handles complex professional work.

Requestygpt-5.4-minigpt-mini

Quick Info

Powered by
Provider
Requesty
Model key
gpt-5.4-mini
Release date
Mar 17, 2026
Last updated
Mar 17, 2026
Knowledge cutoff
2025-08-31
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$4.50

Limits

Input tokens
272,000 tokens
Output tokens
128,000 tokens
Context window
400,000 tokens

Latest news about GPT-5.4 mini

Videos about GPT-5.4 mini

Recent tweets and retweets from Requesty

More models around GPT-5.4 mini