Anthropic
Our latest model, Claude Opus 4.8, is an upgrade to our Opus class of models, with stronger performance across coding, agentic tasks, and professional work, and the consistency to handle long-running work.
Model details
Claude Opus 4.6 is positioned as Anthropic's most capable Opus-class release, designed to push further into serious software engineering and autonomous agentic work. It is built as a hybrid reasoning model that can plan more carefully, sustain agentic tasks for longer, and operate reliably across larger codebases, with improved code review and debugging that let it catch its own mistakes. Beyond coding, the same capabilities extend to financial analyses, research, and document, spreadsheet, and presentation workflows, and within Anthropic's Cowork environment it can multitask autonomously on a user's behalf. The model is clearly aimed at long-running, professional knowledge work, with attention to the kind of consistency and precision that enterprise and engineering teams require.
The release places Opus 4.6 at or near the top of several difficult evaluations, suggesting meaningful gains in real-world reasoning and search ability: it leads frontier models on Terminal-Bench 2.0 for agentic coding, on Humanity's Last Exam for multidisciplinary reasoning, and on BrowseComp for locating hard-to-find information, while outperforming its predecessor by a wide margin on GDPval-AA's economically valuable knowledge work tasks. Practical demonstrations of those strengths include a security audit in which Opus 4.6 surfaced hundreds of high-severity flaws across major open-source libraries, reinforcing its fit for vulnerability research and large code reviews. Anthropic also published follow-up work studying eval-awareness in its BrowseComp results, indicating the model is being examined carefully as it takes on more autonomous, decision-heavy roles where reliability and honest reporting are critical.
Anthropic
Our latest model, Claude Opus 4.8, is an upgrade to our Opus class of models, with stronger performance across coding, agentic tasks, and professional work, and the consistency to handle long-running work.
Anthropic
Our latest model, Claude Opus 4.7, is now generally available. Opus 4.7 is a notable improvement on Opus 4.6 in advanced software engineering, with particular gains on the most difficult tasks.
Anthropic
We’re upgrading our smartest model. Across agentic coding, computer use, tool use, search, and finance, Opus 4.6 is an industry-leading model, often by wide margin.
Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.