Anthropic Details How Claude's AI Watermarks Actually Work
The company explains its SynthID-Text implementation, addressing user concerns about editing, code generation, and detection methods.

Anthropic released new technical details Friday explaining how its Claude chatbot will watermark AI-generated text, responding to user questions and concerns that emerged after the company announced the feature earlier this week.
The watermarking system, required under the EU AI Act's Transparency Code, embeds detectable patterns into Claude's responses without affecting output quality or readability, according to the company's blog post first reported by TechCrunch.
How the watermarking technology works
Anthropic is implementing SynthID-Text, a watermarking approach developed by Google DeepMind in 2024. The system works by creating patterns when Claude makes "low-stakes choices" between equally valid options—such as selecting "overcast" versus "grey" to describe weather.
These patterns remain invisible to readers but can be detected by anyone with the encoding key. Anthropic plans to release a watermark detection API to enable verification.
The company emphasized that its watermarking differs fundamentally from AI detection tools offered by companies like Pangram, which identify AI-generated text by spotting common linguistic patterns. Watermark detection checks for the embedded signal rather than analyzing writing style.
Editing and removal concerns
User concerns about watermark persistence led Anthropic to address editing scenarios directly. Light editing likely won't eliminate the watermark completely, the company said. However, a complete rewrite replacing every word will remove it—though at that point, the text may no longer qualify as AI-generated.
For content where Claude only provided proofreading or light editing assistance, watermark detection depends on text length and the extent of Claude's involvement. When humans write most words themselves, little remains for the watermark to attach to.
Code generation implications
Code presents unique watermarking challenges because the model must prioritize functionality over stylistic choices. Anthropic acknowledged that code will carry less watermarking than other text types, since Claude has limited freedom to vary its output while maintaining working code.
Watermarks can still appear in areas with arbitrary choices, such as code comments, but will have "negligible effect on the actual code produced," according to the company.
Why it matters
The watermarking requirement affects all major AI providers operating in the European Union, not just Anthropic. Other model developers who signed the same Code of Practice will implement their own watermarking systems, creating an industry-wide shift toward traceable AI content. For enterprises using AI tools, this transparency mechanism could help address concerns about undisclosed AI usage in professional contexts, though the technical limitations around code and heavily edited text suggest watermarking won't provide foolproof detection.
The announcement sparked debate among Claude users, with some canceling subscriptions over the change, Business Insider reported. Reddit discussions ranged from conspiracy theories to arguments that only those seeking to deceive would oppose watermarking.
These details were first reported by TechCrunch.
This is an original analysis by the Omega editorial team. Source reporting: AI Watch.
Want systems like this working for your business?
Book a Call
