Currently listed through these providers:
Model details
GPT-5.1
GPT-5.1 continues the GPT-5 series with an emphasis on balancing intelligence and responsiveness for developer workloads, particularly agentic and coding tasks. OpenAI framed the release around adaptive reasoning, where the model dynamically adjusts how long it thinks based on task complexity, becoming faster and more token-efficient on simpler requests while still delivering frontier-level answers when problems demand deeper analysis. A "no reasoning" mode is available for cases where immediate, low-latency responses matter more than deliberation, giving integrators a knob to trade depth for speed. The model also ships with extended prompt caching that retains prompts for up to 24 hours, which helps reduce repeated cost and latency for multi-turn sessions and follow-up questions.
The release leans heavily into practical developer fit rather than benchmark spectacle, prioritizing usability gains like better instruction following, clearer phrasing, and more controllable tone. OpenAI worked directly with coding-tool startups including Cursor, Cognition, Augment Code, Factory, and Warp to refine GPT-5.1's coding personality, steerability, and output quality, and the model debuts alongside an apply patch tool for reliable code edits and a shell tool for executing commands, making it well suited to agentic coding pipelines. Priority Processing tier customers are expected to see noticeably faster performance compared with the prior generation, reinforcing the model's positioning as a responsive daily driver for production applications that mix routine tasks with occasional harder problems.
Quick Info
Powered by- Provider
- Ofox
- Model key
- openai/gpt-5.1
- Release date
- Nov 13, 2025
- Last updated
- Nov 13, 2025
- Knowledge cutoff
- 2024-09-30
- AI SDK package
@ai-sdk/openai- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.25
- Output token cost
- $10.00
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens