GPT-5.3 Codex is OpenAI's latest entry in the Codex line, positioning itself as the most capable agentic coding model released to date. It brings together the frontier software-engineering performance of GPT-5.2 Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2, merged into a single model that OpenAI also reports runs about 25% faster than its predecessor. This unified design aims at long-running work that mixes research, tool use, and multi-step execution rather than short, single-turn coding prompts, making it suitable for end-to-end agentic workflows on a computer.
Beyond raw capability, GPT-5.3 Codex is built for sustained, interactive collaboration. OpenAI highlights that users can steer and interact with the model mid-task without losing context, treating the agent more like a working colleague than a one-shot code generator. The model also sets a new industry high on SWE-Bench Pro and Terminal-Bench, with strong performance on OSWorld and GDPval benchmarks that span coding, agentic execution, and real-world professional tasks. Notably, GPT-5.3 Codex is the first model that OpenAI reports was instrumental in its own creation, as earlier versions helped debug training, manage deployment, and diagnose evaluation results during development.