Anthropic to Watermark All AI-Generated Text From Claude Models Starting August 2
Anthropic confirmed it will watermark AI-generated text and files from Claude models, complying with the EU AI Act's Transparency Code that took effect August 2. The watermark is applied at the model level and persists through copy-paste, according to the company.
Anthropic will watermark text generated by its Claude models to comply with the European Union's AI Act Transparency Code, which took effect on August 2, 2026. The company confirmed the move in an updated support page rather than a formal announcement.
The EU's Transparency Code requires AI companies to mark AI-generated or AI-edited content in a way that other systems can detect. Anthropic said all models released after August 2 will automatically include technology to watermark both computer-generated text and files. For files, the company said it is using the C2PA open standard, an industry specification already adopted by other AI and media companies for content provenance.
Anthropic also said it plans to extend watermarking support to older models, though it did not specify a timeline. According to the company's support page, the watermark is embedded directly in the generated text: "Because the watermark is part of the text, it will travel with the text when it's copied and pasted elsewhere, and may persist through some editing. Watermarking will be applied at the model level, which means it will be present no matter which Claude product or surface the text comes from."
The watermarking will apply across Anthropic's product line, including the Claude platform API, the Claude consumer app, Claude Code, Claude Cowork, and Claude Tag.
One open question: how much a user needs to edit AI-generated text before the watermark breaks or disappears. Anthropic has not disclosed the technical threshold, and TechCrunch has asked the company for clarification.
Anthropic joins a growing list of companies that have committed to the EU's Transparency Code, including Black Forest Labs, Google, Meta, Microsoft, OpenAI, and Synthesia. The push toward watermarking has accelerated across the industry in recent weeks. AI music platform Suno said last week it would mark tracks generated on its platform following legal challenges over copyright. Newsletter platform Substack partnered with detection firm Pangram last month to flag AI-generated content, with CEO Chris Best coining the term "Claudefishing" to describe writers passing off AI-generated text as their own.
What this means
This is a compliance move, not a voluntary transparency push. The EU AI Act's Transparency Code gives companies a legal deadline, and Anthropic's quiet support-page update — rather than a press release — suggests the company is treating this as a regulatory checkbox rather than a marketing opportunity.
The unresolved question about editing thresholds matters more than it might seem. If light edits strip the watermark, the system offers minimal real-world detection value against anyone motivated to evade it — students avoiding academic detection, or content farms avoiding platform bans. Watermarking that survives copy-paste but not paraphrasing catches casual users, not determined ones.
Expect this to become the industry baseline rather than a differentiator. With Google, Meta, Microsoft, and OpenAI already committed to the same EU code, watermarking is turning into table-stakes regulatory compliance for any AI company operating in Europe, not a competitive feature. The real test will be whether these watermarks hold up against third-party detection tools and adversarial editing in practice.
Related Articles
Anthropic Cuts False Positives in Fable 5's Biology Filter by 85%, Keeps Virology and Toxicology Blocked
Anthropic has cut false positives in Fable 5's biology safety classifier by roughly 85%, letting users ask about lab results, symptoms, and medical questions without being rerouted to the weaker Opus 5 model. Dual-use topics like virology, toxicology, and molecular design remain restricted, with Anthropic citing the difficulty of containing biological threats once released.
Anthropic SDK v0.121.0 Adds Session Budgets, Mid-Conversation Tool Changes, and GitHub Skills Auto-Loading
Anthropic released version 0.121.0 of its Python SDK on August 7, 2026, introducing a new beta for mid-conversation tool changes, session budgets, an advisor tool, pinned inference location, and skills auto-loading from GitHub. The update also removes retired Claude Opus 4.1 models from the API.
Anthropic Makes Claude Code's Auto Mode Default for Pro, Max, and Team Users on August 14
Anthropic will make Claude Code's auto mode the default for Pro, Max, and Team accounts starting August 14, reducing step-by-step approval prompts. The company cites a study of 1,053 testers showing auto mode caught 89% of harmful actions versus 13.6% for manual review.
Anthropic Makes Auto Mode Default in Claude Code for Pro, Max, and Team Plans Starting August 14
Anthropic will make auto mode the default setting for new Claude Code sessions on Pro, Max, and Team plans starting August 14, 2026. The company cites a 1,053-person study showing auto mode blocked 89% of harmful actions compared to 13.6% for human reviewers, plus a third-party test claiming zero successful prompt injections out of 720 attempts.
Comments
Loading...