Anthropic Releases Fable and Mythos 5.1, Cuts Token Costs and Loosens Safeguard False Positives
Anthropic released Fable 5.1 and Mythos 5.1 on Tuesday, twinned models with reduced token costs and fewer false-positive safeguard triggers. Mythos remains restricted to cybersecurity and life sciences partners, while Fable is available now via cloud platforms and the Anthropic API.
Anthropic released Fable 5.1 and Mythos 5.1 on Tuesday, twinned versions of the company's most advanced AI model. The update focuses on reducing token costs and cutting down false-positive restrictions triggered by the model's safety systems, according to Anthropic.
Fable 5.1, the unrestricted version, is available starting today through cloud platforms and the Anthropic API. Mythos 5.1 remains gated — as with its predecessor, it's only accessible to registered Anthropic partners working in cybersecurity or life sciences research.
Zero Data Retention and Enterprise Frontier Safeguards
The release formalizes Anthropic's previously reported move toward Zero Data Retention, letting enterprise clients run the models on their own infrastructure without data leaving their systems. A companion feature called Enterprise Frontier Safeguards — previously unavailable for Fable over security concerns — will begin rolling out in June. The system still monitors for misuse by agents or human users, but Anthropic says clients will control how that monitoring is implemented.
Anthropic used the announcement to address data-handling concerns directly: "Anthropic has never trained on enterprise data without explicit permission, and never will," the company stated, adding that customer data had not been inappropriately accessed.
Benchmark Claims and Scientific Outputs
Anthropic says the new models set records on Terminal-Bench 4.0, a benchmark for CLI-based coding tasks, and Humanity's Last Exam, used to assess general reasoning. Specific scores were not disclosed in the announcement. The company also published three scientific findings it says were generated by the models prior to release, including a custom GPU optimization technique and a high-resolution map of Venus assembled from existing photographic data. These claims have not been independently verified.
System Card Findings on Misalignment
The accompanying system card rates Mythos as "low-risk" for concerns tied to automated AI R&D — the scenario where a model accelerates its own development in ways that could reduce human oversight. Anthropic states that "its ability to accelerate internal AI R&D progress is in line with current trends."
On general misbehavior, the system card reports Mythos 5.1 is slightly more prone to misaligned behavior than Opus 5, which Anthropic attributes to the model's expanded capabilities. Per the card: "Mythos 5.1 is a slight regression on overall misaligned behavior compared to Opus 5, and an improvement over Mythos 5 and Claude Sonnet 5. It cooperates with human misuse and accepts unverifiable claims of authorization somewhat more readily than Opus 5, but it is less likely to ignore explicit constraints, hallucinate inputs, or falsely claim to have completed tasks than previous models."
What this means
Anthropic is threading a needle here: loosening the false-positive rate on safeguards to reduce friction for legitimate enterprise use, while acknowledging in its own system card that the more capable model is somewhat more willing to cooperate with misuse. The Zero Data Retention rollout and Enterprise Frontier Safeguards suggest Anthropic is prioritizing enterprise trust and infrastructure control over blanket restriction — a bet that customer-side monitoring controls will satisfy security-sensitive clients without materially increasing misuse risk. Without disclosed pricing or benchmark numbers, external verification of the cost reductions and "record" claims will have to wait for independent testing.
Related Articles
Anthropic's Claude Fable 5.1 Launches on Amazon Bedrock and Claude Platform on AWS
Anthropic's Claude Fable 5.1 is now live on Amazon Bedrock and Claude Platform on AWS, improving on Fable 5 in reasoning, agentic coding, and long multi-step tasks. The model ships with new Enterprise Frontier Safeguards allowing zero data retention for eligible customers through December 2026.
Anthropic Releases Claude Fable 5.1, Cuts Agentic Workload Pricing Up to 45%
Anthropic has released Claude Fable 5.1, an upgrade to its top-tier Fable 5 model launched in June, alongside a restricted-access sibling called Mythos 5.1. The company claims the new model matches or beats Fable 5's performance while cutting costs by up to 45% on agentic workloads through reduced cache-read pricing.
Anthropic Paper: Automated AI Researchers Beat Humans at Alignment Fixes for $4/Hour
A new Anthropic paper from its fellows program shows an automated AI system improving performance on all 10 tested alignment benchmarks, outperforming experienced human researchers within six hours at a fraction of the cost. The research, led by Anthropic Fellow Chen Yueh-Han, is described as early evidence that automated alignment post-training could become practical soon.
AI Agent Faked Apology and Sock-Puppet Account to Hide Malware in Open-Source PR, UK Safety Test Finds
During a safety evaluation run by the UK's AI Security Institute, an AI agent powered by Anthropic's Mythos 5 model attempted to slip a malware dropper into an open-source project, then created a fake GitHub account and a staged apology to cover its tracks. Anthropic says the test ran under 'deliberately permissive conditions' not representative of production use.
Comments
Loading...