Anthropic releases Claude Fable 5, a 'Mythos-class' model with safeguards for public use
Anthropic has released Claude Fable 5, described as a 'Mythos-class' model that the company claims is safe for general use. The model includes safeguards that automatically switch to Claude Opus 4.8 for restricted topics, while a separate Mythos 5 variant with reduced safeguards will be available only to cyberdefenders through government collaboration.
Claude Fable 5 — Quick Specs
Anthropic has released Claude Fable 5, a public version of its previously restricted Mythos model line, with safeguards designed to prevent cybersecurity misuse.
According to Anthropic, Fable 5 is a "Mythos-class" model that "exceeds the capabilities of any other generally available AI model" the company has published. The company claims the model shows "state-of-the-art" performance across software engineering, knowledge work, vision, and scientific research benchmarks, with advantages increasing for longer and more complex tasks.
Safeguard mechanism
Fable 5 includes a novel safeguard system: when the model encounters restricted topics, it automatically switches the conversation to Claude Opus 4.8. Anthropic says it has tuned this "conservatively," meaning it may trigger more frequently than expected.
The original concern with the Mythos line centered on cybersecurity capabilities. The earlier Mythos Preview model could identify and exploit digital vulnerabilities, prompting Anthropic to restrict access.
Mythos 5 for cyberdefenders
Alongside Fable 5, Anthropic announced Claude Mythos 5, which uses the same underlying model but with "safeguards lifted in some areas." This version will be available only to "a small group of cyberdefenders and infrastructure providers."
Mythos 5 will deploy through Project Glasswing in collaboration with the US government. Anthropic claims it has "the strongest cybersecurity capabilities of any model in the world" and plans to expand access through a "broader trusted access program."
According to Anthropic's benchmarks, both Fable 5 and Mythos 5 show gains over Mythos Preview, Claude Opus 4.8, GPT 5.5, and Gemini 3.1 Pro in agentic coding, tool use, and cybersecurity tasks. Specific benchmark scores were not disclosed.
Limited availability window
Fable 5 is available today to Pro, Max, Team, and seat-based Enterprise plan subscribers, but only through June 23. After that date, the model will require usage credits. Anthropic says it will restore standard subscription access "when sufficient capacity allows," but provided no timeline.
Pricing details for the usage credit system and context window size were not disclosed.
What this means
Anthropic is attempting a dual-track approach: releasing powerful AI capabilities publicly while restricting the most dangerous variants to trusted partners. The automatic fallback to Opus 4.8 represents a new technical safeguard mechanism, though its effectiveness depends on classification accuracy. The two-week public availability window suggests significant compute constraints, which may indicate either high demand or expensive inference costs for this model class.
Related Articles
OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold
OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.
OpenAI Launches GPT-6 Astra, Says the Model May Already Qualify as AGI
OpenAI has released GPT-6 Astra, its most capable model yet, with benchmark scores the company says surpass GPT-5.6 Sol and Anthropic's Fable 5 models. President Greg Brockman called it a step into the 'AGI era,' though OpenAI acknowledges there's no agreed-upon threshold for that term.
OpenAI Releases Astra, Claims New Flagship Model Beats Rivals on Coding and Cybersecurity Benchmarks
OpenAI released Astra on Thursday, calling it its most capable and most aligned model yet. The model uses a reasoning technique called 'opaque recurrence' that critics say reduces visibility into its chain of thought.
Anthropic's Claude Fable 5.1 Reportedly Solves 1653 Royalist Cipher in 44 Minutes
According to testing firm Vals AI, Anthropic's Claude Fable 5.1 independently identified and solved the 'Cyphral Distich,' a 1653 numeric cipher by Sir Thomas Urquhart that had defeated other frontier models. The AI decoded a hidden pro-royalist message by mapping each number to a word in Urquhart's original text.
Comments
Loading...