Sulat.com
AI models
302.AI logo

Model details

gpt-5.4-mini

GPT-5.4 mini sits inside the gpt-mini family as a compact sibling to the larger GPT-5.4, designed to bring much of that flagship's capability to latency-sensitive, high-volume workflows. OpenAI positioned it as a meaningful upgrade over GPT-5 mini across coding, reasoning, multimodal understanding, and tool use, while running more than twice as fast, and on evaluations such as SWE-Bench Pro and OSWorld-Verified it approaches the scores of the full-size GPT-5.4 model. That balance of speed and near-flagship reasoning makes it well suited to responsive coding assistants, subagents handling supporting tasks, computer-use systems that interpret screenshots, and multimodal applications that need to reason about images in real time, where the Decoder also notes it competes in the same efficiency tier as Gemini 3 Flash.

In practical terms, the model combines a very large context window with strong agentic tooling, supporting both text and image inputs and producing text outputs with capabilities for attachments, reasoning, tool calling, and structured outputs. This makes it a flexible drop-in for pipelines that need reliable function calling and schema-bound responses alongside genuine multimodal comprehension, rather than a narrow single-purpose model. Independent coverage frames it as part of OpenAI's broader push into cost-tiered variants, offering a faster, more capable option for teams that previously had to choose between mini-class efficiency and flagship-class quality.

302.AIgpt-5.4-minigpt-mini

Quick Info

Powered by
Provider
302.AI
Model key
gpt-5.4-mini
Release date
Mar 19, 2026
Last updated
Mar 19, 2026
Knowledge cutoff
2025-08-31
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.75
Output token cost
$4.50

Limits

Input tokens
272,000 tokens
Output tokens
128,000 tokens
Context window
400,000 tokens

Transparent token rates

Compare gpt-5.4-mini pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about gpt-5.4-mini

302.AI

CoverageRelease Notes

OpenAI's official Model Release Notes confirm that GPT-5.4 mini began rolling out in ChatGPT on March 18, 2026, available to Free and Go users via the "Thinking" feature in the + menu, and serving as a rate-limit fallback for GPT-5.4 Thinking on Plus, Pro, and other paid tiers; Enterprise customers retain the option to The same release-notes page shows the GPT-5.4 mini entry sitting between later entries for GPT-5.5 Instant (May 28, 2026) and GPT-5.6 Sol (July 9, 2026), framing gpt-5.4-mini as a mid-generation, fallback-tier model rather than a flagship. For 302.AI API consumers, the practical implication is that gpt-5.4-mini remains

302.AI

Coverage

OpenAI has officially bridged the gap in its model lineup with the debut of GPT-5.4 Mini and GPT-5.4 Nano. These models are designed to compete with high-efficiency models like Gemini 3 Flash., OpenAI has officially bridged the gap in its model lineup with the debut of GPT-5.4 Mini and GPT-5.4 Nano. These models are de

302.AI

Coverage

Tech News News: OpenAI has launched two new AI models — GPT-5.4 mini and GPT-5.4 nano — aimed at delivering faster performance and lower costs for high-volume workloa.

302.AI

Coverage

OpenAI has released two new compact models—GPT-5.4 mini and nano—built for coding assistants, subagents, and computer control. GPT-5.4 mini nearly matches the full model's performance, but both new models come with a steep price hike over their predecessors.

302.AI

Coverage

The Sulat model catalog entry for 302.AI's gpt-5.4-mini-2026-03-17 confirms it as a mid-tier variant in OpenAI's GPT-5 family, released on March 18, 2026, with a knowledge cutoff of August 31, 2025. It documents a 272,000-token input limit, a 128,000-token maximum output, and a combined 400,000-token context window, al On the OpenAI-compatible 302.AI gateway, the model is priced at $0.75 per million input tokens and $4.50 per million output tokens, positioning it as a cost-effective default for high-volume, low-latency production workloads. The listing also notes that the previous GPT-5 mini remains relevant mainly for legacy workloa

Videos about gpt-5.4-mini

More models around gpt-5.4-mini