Sulat.com
AI models
DevPass (LLM Gateway) logo

Model details

GPT-5.4 nano

GPT-5.4 nano is the smallest and most economical tier in OpenAI's GPT-5.4 lineup, announced on March 17, 2026 alongside GPT-5.4 mini as a fast, efficient alternative to the flagship model. Press coverage describes the model as inheriting many of the strengths of the larger GPT-5.4 while being tuned for coding assistance, subagent orchestration, and other high-volume workloads where reduced cost and quick turnaround matter more than peak capability. Its multimodal design accepts both text and image inputs while producing text-only outputs, giving it enough flexibility for grounded chat, document analysis, and tool-mediated pipelines without the overhead of a larger model.

For developers routing requests through the LLM Gateway, GPT-5.4 nano surfaces meaningfully different behavior depending on the upstream provider. OpenRouter's per-provider telemetry shows OpenAI direct delivering the fastest response, with a P50 latency of 0.64 seconds at roughly 59 tokens per second, while Azure and Azure US trade speed for higher availability at 99.75% and 99.95% uptime respectively, and an OpenAI Flex tier offers a separate throughput profile. That spread makes the model well suited to production scenarios where teams can choose between raw speed and steady reliability, particularly for embedded assistants, batch summarization, and lightweight agent loops where the the cataloged API limit context window comfortably accommodates long conversational histories or document collections.

DevPass (LLM Gateway)gpt-5.4-nanogpt-nano

Quick Info

Powered by
Provider
DevPass (LLM Gateway)
Model key
gpt-5.4-nano
Release date
Mar 17, 2026
Last updated
Mar 17, 2026
Knowledge cutoff
2025-08-31
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.20
Output token cost
$1.25

Limits

Input tokens
272,000 tokens
Output tokens
128,000 tokens
Context window
400,000 tokens

Transparent token rates

Compare gpt-nano pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GPT-5.4 nano

LLM Gateway

Coverage

Tech News News: OpenAI has launched two new AI models — GPT-5.4 mini and GPT-5.4 nano — aimed at delivering faster performance and lower costs for high-volume workloa.

LLM Gateway

Coverage

OpenAI has launched GPT-5.4 mini and nano, two smaller models optimized for coding, subagents, and high-volume workloads at reduced cost.

LLM Gateway

Coverage

On March 17, 2026, OpenAI announced the release of ' GPT-5.4 mini ' and ' GPT-5.4 nano ,' lightweight versions of GPT-5.4, which debuted in March 2026. GPT-5.4 mini/nano are designed to be fast and efficient models that can handle high processing loads while inheriting many of the strengths of GPT-5.4. OpenAI introduce

DevPass (LLM Gateway)

CoverageBenchmark

OpenRouter's listing for OpenAI's GPT-5.4 Nano confirms the model as the lightweight, cost-optimized member of the GPT-5.4 family, released March 17, 2026 with a 400K context window and text-plus-image input modalities. Listed pricing is $0.20 per 1M input tokens and $1.25 per 1M output tokens, with OpenAI Flex offerin For developers routing through the LLM Gateway, the OpenRouter page provides actionable per-provider performance snapshots: OpenAI direct reports P50 latency of 0.64s at ~59 tps with 90.25% uptime, Azure shows 1.42s at 24 tps with 99.75% uptime, Azure (US) shows 1.66s at 30 tps with 99.95% uptime, and OpenAI Flex track

Videos about GPT-5.4 nano

More models around GPT-5.4 nano