Poe
Claude Opus 4.7 hits 87.6% on SWE-bench Verified, adds /ultrareview, xhigh mode, and task budgets. Full benchmarks, pricing, and 4.6 migration guide.
Model details
Claude Opus 4.7 represents the latest evolution in Anthropic's premium flagship tier, positioned as the most capable generally available model for complex enterprise and agentic work. The model builds on its predecessor's architecture to deliver particularly strong gains in advanced software engineering, where users report being able to hand off their most difficult coding tasks with confidence. The architecture emphasizes rigorous, consistent execution on long-running tasks with precise attention to instructions, including mechanisms for the model to verify its own outputs before reporting back. Beyond pure coding, the model brings improved taste and creativity to professional tasks, producing higher-quality interfaces, slides, and documentation. The vision system also sees images in greater resolution, enabling stronger performance on document work, chart analysis, and visual verification tasks across extended workflows.
The training lineage shows a deliberate refinement of the Opus family, with improvements grounded in real production use cases rather than benchmark optimization alone. Opus 4.7 was developed with a measured approach to capability scaling, explicitly positioned below the more powerful Claude Mythos Preview in raw capability while offering better results than Opus 4.6 across tested benchmarks. The model was the first to receive new cybersecurity safeguards developed by Anthropic after the Mythos Preview announcement, with targeted efforts during training to differentially reduce certain capabilities. The focus on sustained autonomy and execution across extended sessions makes it especially effective for asynchronous agent pipelines—large codebases, multi-stage debugging, and end-to-end project orchestration—where persistence, judgment, and follow-through matter more than answering a single question.
Poe
Claude Opus 4.7 hits 87.6% on SWE-bench Verified, adds /ultrareview, xhigh mode, and task budgets. Full benchmarks, pricing, and 4.6 migration guide.