Currently listed through these providers:
Model details
GPT-5.1 Codex Max
GPT-5.1-Codex-Max represents a fundamental shift in what AI coding assistants can handle. Unlike earlier iterations that offered code completion and chat-based suggestions, this model was built from the ground up for autonomous development work across massive codebases. Its defining innovation is context compaction technology, which lets it coherently operate across millions of tokens in a single task—working through entire features, implementation, and testing cycles without losing track of the bigger picture. The architecture builds on an updated version of the 5.1 reasoning stack and introduces multiple reasoning effort levels, from none through to a new xhigh tier designed for maximum code quality on the most complex problems.
The model delivers measurable improvements in both speed and token efficiency, achieving 77.9% on SWE-bench Verified with 30% fewer thinking tokens than previous approaches, and scores 58.1% on Terminal Bench 2.0 compared to competing models' lower results. OpenAI has tested the model's limits by observing it work continuously for over 24 hours, autonomously iterating through code and fixing test failures without human intervention—demonstrating the kind of persistence that production development workflows demand. This combination of long-horizon reasoning, proven benchmark performance, and extended autonomous operation makes it particularly suited for agentic coding pipelines where sustained, multi-stage development work replaces one-off completions.
Quick Info
Powered by- Provider
- OpenCode Zen
- Model key
- gpt-5.1-codex-max
- Release date
- Nov 13, 2025
- Last updated
- Nov 13, 2025
- Knowledge cutoff
- 2024-09-30
- AI SDK package
@ai-sdk/openai- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.25
- Output token cost
- $10.00
Limits
- Input tokens
- 272,000 tokens
- Output tokens
- 128,000 tokens
- Context window
- 400,000 tokens
Transparent token rates
Compare gpt-codex pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GPT-5.1 Codex Max
Videos about GPT-5.1 Codex Max
More models around GPT-5.1 Codex Max
This exact model name is also listed by 10 other providers.