Model details
GPT-5.5
GPT-5.5 is OpenAI's first fully retrained base model since GPT-4.5, internally codenamed "Spud," and represents a ground-up rebuild rather than another incremental update on the prior architectural foundation. According to third-party analysis, it moves away from stitched-together modality handlers and instead processes text, images, audio, and video end-to-end within a single unified architecture. That omnimodal redesign is the headline structural change of the release, intended to let one model reason across formats without passing inputs between specialized submodels.
Beyond architecture, the release is framed around tighter hardware and infrastructure integration. The model was reportedly co-designed alongside NVIDIA's GB200-class rack-scale systems, which third-party reporting credits with helping GPT-5.5 match the per-token latency of its predecessor despite a clear capability jump. OpenAI's own Codex tooling was also used to rewrite portions of the serving stack before launch, reportedly yielding faster token generation once the model went live. Taken together, the picture is a flagship aimed at long-context, multi-format assistant workloads, where unified multimodal understanding, sustained reasoning, and high-throughput generation matter more than any single benchmark number.
Quick Info
Powered by- Provider
- OpenAI
- Model key
- gpt-5.5
- Release date
- Apr 23, 2026
- Last updated
- Apr 23, 2026
- Knowledge cutoff
- 2025-12-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $5.00
- Output token cost
- $30.00
Limits
- Input tokens
- 922,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 1,050,000 tokens
OpenCode
Model variants
Transparent token rates
Compare GPT-5.5 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GPT-5.5
OpenAI
OpenAI's GPT-5.5 boosts agentic coding, reduces costs, and handles complex tasks with minimal input across business and research use.
OpenAI
OpenRouter's model page for OpenAI's GPT-5.5 lists the model with a 1M+ token context window (922K input / 128K output), text and image input modalities, a knowledge cutoff of December 2025, and a release date of April 24, 2026. Standard list pricing is shown as $5 per 1M input tokens and $30 per 1M output tokens, with Provider-level performance metrics on the page indicate a P50 latency of 3.16 seconds and throughput of about 46 tokens per second on the OpenAI route, with uptime ranging from 98.76% to 100% across providers. Weighted-average effective input pricing after prompt caching is reported around $1.42–$1.76 per 1M tokens, re
Videos about GPT-5.5
More models around GPT-5.5
This exact model name is also listed by 40 other providers.