model releaseAnthropic

Anthropic Restores Claude Fable 5 After Government Takedown, With Stricter Cybersecurity Blocks

TL;DR

Anthropic is redeploying Claude Fable 5 after a month-long government-mandated takedown triggered by Amazon researchers discovering a method to bypass the model's cybersecurity safeguards. The returning version includes enhanced safety classifiers that automatically block cybersecurity tasks and revert to Opus 4.8, with restricted availability through usage credits only.

2 min read
0

Claude Fable 5 — Quick Specs

Context window1000K tokens
Input$10/1M tokens
Output$50/1M tokens

Anthropic Restores Claude Fable 5 After Government Takedown, With Stricter Cybersecurity Blocks

Anthropic is redeploying Claude Fable 5 on July 1, approximately one month after the US government ordered the model's removal from public access. The returning version includes what researchers describe as "extraordinarily strong" safeguards that automatically block cybersecurity-related tasks.

Why Fable 5 Was Taken Down

According to Anthropic, Amazon researchers discovered a method to bypass Fable 5's original safeguards and reported the vulnerability to the US government. The testing involved prompting the model to identify software weaknesses, which was classified as a high-security task.

Anthropic claims its own testing found that less capable models—including Claude Opus 4.8, GPT-5.5, and Kimi K2.7—could identify the same vulnerabilities. The company states that every model tested could produce the same exploit demonstrations, including Claude Haiku 4.5, Sonnet 4.6, Opus 4.6, Opus 4.7, Opus 4.8, GPT-5.4, GPT-5.5, and Kimi K2.7.

How the New Version Works

The redeployed Fable 5 features an improved safety classifier trained in collaboration with the US government. When the model detects a potentially high-risk task, it will automatically block the request and redirect it to Opus 4.8 instead. Users will receive a notification when this occurs.

Anthropic warns that this switching behavior may trigger during routine tasks like coding and debugging—not because Fable 5 lacks the capability, but due to the imposed safeguards. The company acknowledges this represents a stricter implementation than the original release, though it states "this might not be the case for 99% of tasks."

Restricted Availability

Fable 5 will not be freely accessible through standard usage limits. From July 1-7, Pro, Max, Team, and select Enterprise plans will have access using 50% of their usage limit. After July 7, the model will only be available via usage credits.

The model consumes significantly more tokens than standard Claude models, eating through usage limits faster. Anthropic positions Fable 5 and its cybersecurity-focused counterpart Mythos 5 as designed for complex tasks rather than routine chatbot interactions.

What This Means

The Fable 5 incident marks one of the first cases of a major AI model being temporarily banned by government order over security concerns. Anthropic's response—implementing automatic task-blocking that reverts to a less capable model—sets a precedent for how AI companies may handle government pressure on advanced models.

However, Anthropic's own testing suggesting that less capable models could perform the same exploits raises questions about whether the restrictions meaningfully improve security or simply create operational friction. The company's claim that the vulnerability "could have been done with any other model" undermines the rationale for Fable 5's specific targeting.

Related Articles

research

Anthropic's Claude Fable 5.1 Reportedly Solves 1653 Royalist Cipher in 44 Minutes

According to testing firm Vals AI, Anthropic's Claude Fable 5.1 independently identified and solved the 'Cyphral Distich,' a 1653 numeric cipher by Sir Thomas Urquhart that had defeated other frontier models. The AI decoded a hidden pro-royalist message by mapping each number to a word in Urquhart's original text.

product update

Anthropic Brings Background Computer Use to Claude Code and Cowork on Mac

Anthropic has enabled background computer use for Claude Code and Claude Cowork on macOS, available to Pro and Max subscribers. The feature lets Claude click, type, and open apps on a Mac without taking over the user's active cursor, following a similar launch by OpenAI's ChatGPT earlier in 2026.

model release

OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold

OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.

model release

OpenAI Launches GPT-6 Astra, Says the Model May Already Qualify as AGI

OpenAI has released GPT-6 Astra, its most capable model yet, with benchmark scores the company says surpass GPT-5.6 Sol and Anthropic's Fable 5 models. President Greg Brockman called it a step into the 'AGI era,' though OpenAI acknowledges there's no agreed-upon threshold for that term.

Comments

Loading...