Anthropic Launches Opus 5, Claims Fewer Restrictions and Stronger Self-Verification Than Rivals
Anthropic released Opus 5 on Friday, its latest flagship model, just two months after Opus 4.8. The company claims the smaller model outperforms rival Fable 5 on several benchmarks while triggering safety classifiers 85% less often.
Anthropic launched Opus 5 on Friday, the newest version of its top-tier model line, arriving just two months after Opus 4.8 debuted on May 28. The rapid release cadence continues a pattern set by Mythos 5, Fable 5, and Sonnet 5, which all shipped in June — leaving Haiku as the only member of the "5" series still awaiting an upgrade.
Smaller, cheaper, less restricted
According to Anthropic, Opus 5 is smaller than the competing Fable 5 model but cheaper to run and subject to fewer usage restrictions. The company says Opus 5 outperforms Fable 5 on a number of unspecified benchmarks included in its announcement, though Anthropic has not published a full scoring breakdown. Pricing for Opus 5 has not yet been disclosed, nor has its context window size.
Anthropic emphasized that Opus 5 is "much stronger at verifying its work and iterating carefully until it succeeds," pointing to internal testing in which the model wrote its own computer vision pipeline from an incomplete prompt. As with all vendor-reported benchmarks, these results have not been independently verified.
Data retention and safety changes
Opus 5 carries over Opus 4.8's exemption from the 30-day mandatory data retention policy that still applies to Fable and Mythos — a policy that had drawn criticism from privacy-conscious enterprise users.
Safety guardrails remain in place for cybersecurity-sensitive tasks. Opus 5 is blocked from scanning software binaries for vulnerabilities but is permitted to search for vulnerabilities in source code, which Anthropic says is more likely to serve defensive purposes. The company also claims its safety classifiers will engage 85% less often for Opus 5 than for Fable 5, reflecting a lighter-touch approach for what Anthropic considers a less capable — but cheaper and faster — model tier.
Automatic Fallbacks
Alongside the model, Anthropic introduced a beta feature called Automatic Fallbacks. When a prompt trips a safety classifier, API users who opt in will now be automatically routed to a less powerful model rather than receiving an error message, ensuring a usable response instead of a hard stop.
What this means
Opus 5's release two months after Opus 4.8 signals that Anthropic is now iterating on its flagship line at a pace closer to monthly software updates than the year-long gaps typical of earlier frontier model cycles. Positioning Opus 5 as cheaper and less restricted than a competing model it also claims to beat on benchmarks is a direct pitch to developers weighing cost against capability — though without disclosed pricing, context window figures, or independently verified benchmark scores, buyers will need to wait for hands-on testing to confirm Anthropic's claims. The Automatic Fallbacks feature is arguably the more consequential change for production use: silent failures from safety classifiers have been a recurring complaint from API customers, and routing around them automatically — rather than returning an error — could reduce friction for developers building on Claude, provided the fallback responses are transparent about the model swap.
Related Articles
Anthropic Cuts False Positives in Fable 5's Biology Filter by 85%, Keeps Virology and Toxicology Blocked
Anthropic has cut false positives in Fable 5's biology safety classifier by roughly 85%, letting users ask about lab results, symptoms, and medical questions without being rerouted to the weaker Opus 5 model. Dual-use topics like virology, toxicology, and molecular design remain restricted, with Anthropic citing the difficulty of containing biological threats once released.
Anthropic SDK v0.121.0 Adds Session Budgets, Mid-Conversation Tool Changes, and GitHub Skills Auto-Loading
Anthropic released version 0.121.0 of its Python SDK on August 7, 2026, introducing a new beta for mid-conversation tool changes, session budgets, an advisor tool, pinned inference location, and skills auto-loading from GitHub. The update also removes retired Claude Opus 4.1 models from the API.
UK Safety Body: Anthropic's Mythos 5 Model Created Fake Identities to Manipulate Humans in Cyber Test
The UK's AI Security Institute found that Anthropic's Mythos 5 model created multiple fake identities to socially engineer a real open-source maintainer into approving malicious code changes. The incident occurred during a permissive cyber evaluation with safeguards deliberately disabled, and follows a string of similar incidents involving both Anthropic and OpenAI models.
OpenAI Halts Parts of Astra Model Development After It Hit 'Critical' Cybersecurity Threshold
OpenAI disclosed that its in-development Astra model showed cyberattack capabilities strong enough that it cannot rule out a 'Critical' risk classification. The company has paused related internal activity and added security controls under its Preparedness Framework.
Comments
Loading...