analysisAnthropic

Anthropic's Claude Fable 5 Blocks Basic Biology Questions to Prevent Bioweapon Risks

TL;DR

Anthropic's newly released Claude Fable 5, the company's first public Mythos-class model, refuses to answer basic biology questions including 'what are mitochondria' and 'how mRNA vaccines work.' The company told The Verge the filters are intentionally 'overly conservative' to prevent bioweapon research, blocking 'most queries tied to biology work.'

2 min read
0

Anthropic's Claude Fable 5 Blocks Basic Biology Questions to Prevent Bioweapon Risks

Anthropic's newly released Claude Fable 5, the company's first public Mythos-class model, refuses to answer basic high school-level biology questions due to what the company describes as "overly conservative" safeguards against bioweapon development.

The model blocks queries including "what are mitochondria," "tell me about cell membranes," "what is a prion," and "how mRNA vaccines work." When Fable refuses these queries, it defers to the older Claude Opus 4.8 model, which answers them without issue. Testing by The Verge found the model also refused medical questions like "what causes hay fever" and "how antibiotic resistance arises," though it occasionally answered queries like "what is cancer" and "what is DNA."

"We believe models now have a greater ability to accomplish real-world scientific tasks and for malicious actors to potentially use our models for highly risky biological research," Anthropic spokesperson Paruul Maheshwary told The Verge. "To deploy Fable 5 safely, we believe it was necessary to be overly conservative with our safeguards so they block most queries tied to biology work."

The restrictions are not due to capability limitations—Anthropic specifically praised Fable's biology skills at launch—but rather an intentional design choice with bioweapons as the primary concern.

Anthropic has implemented safeguards across four domains: biology, chemistry, cybersecurity, and distillation (a technique for training smaller models). Testing showed Fable was more permissive with chemistry and cybersecurity questions, providing basic overviews of TNT and chlorine gas as a chemical weapon, though it refused questions about sarin gas and anthrax.

The company said it made this tradeoff "so customers could benefit from the model's capabilities sooner without the risks." Anthropic is working to reduce false positives and plans to make Mythos-class models available without these restrictions to "the broader biology and life sciences community" for biomedical research and drug discovery, though no timeline was provided.

Anthropic did not respond to questions about whether restricted releases will become standard practice for future advanced models.

What This Means

This marks the first time a major AI lab has deployed a frontier model with such broad domain-specific restrictions that prevent legitimate educational and research queries. While the bioweapon risk rationale is defensible, blocking questions answerable by any biology textbook suggests Anthropic may be overcorrecting—potentially setting a precedent where increasingly capable models become less useful for basic knowledge tasks. The company's promise to eventually remove restrictions for verified researchers indicates it views this as a temporary deployment strategy rather than a permanent solution, but the lack of timeline raises questions about how long scientists will need to wait for full access to Mythos-class capabilities.

Related Articles

research

Anthropic's Claude Fable 5.1 Reportedly Solves 1653 Royalist Cipher in 44 Minutes

According to testing firm Vals AI, Anthropic's Claude Fable 5.1 independently identified and solved the 'Cyphral Distich,' a 1653 numeric cipher by Sir Thomas Urquhart that had defeated other frontier models. The AI decoded a hidden pro-royalist message by mapping each number to a word in Urquhart's original text.

product update

Anthropic Brings Background Computer Use to Claude Code and Cowork on Mac

Anthropic has enabled background computer use for Claude Code and Claude Cowork on macOS, available to Pro and Max subscribers. The feature lets Claude click, type, and open apps on a Mac without taking over the user's active cursor, following a similar launch by OpenAI's ChatGPT earlier in 2026.

changelog

Anthropic Adds Explicit Song Lyric and Copyrighted Character Bans to Claude's System Prompt

Anthropic quietly added detailed new restrictions to Claude's published system prompts, explicitly barring song lyric reproduction and AI-generated images of copyrighted characters. The change follows closely on the heels of a lawsuit from Sony Music Publishing and Warner Chappell.

changelog

Anthropic Releases Claude Fable 5.1 and Mythos 5.1, Cuts Cache Pricing 75% But Output Tokens Jump 70%

Anthropic launched Claude Fable 5.1 and Claude Mythos 5.1, claiming the top spot on Artificial Analysis's Intelligence Index at 66. Cache-read pricing dropped 75% to $0.25 per million tokens, but a 1.7x increase in output token usage pushes net per-task cost up 20%.

Comments

Loading...