How Anthropic Implements AI Text Watermarking
Anthropic has announced a new watermarking technique for its Claude AI models, designed to align with the European Union's AI Act. This measure requires AI-generated content to be marked for transparency. The watermarking method is based on SynthID-Text, a technique developed by Google DeepMind and detailed in a 2024 Nature paper. This approach involves using word choice patterns, without adding hidden characters, to embed the watermark.
The watermark is embedded in the text generation process, using randomness derived from a secret key in word selection. This ensures that the text remains indistinguishable to readers, but detectable by those with the key. The watermark does not alter the quality, cost, or speed of the AI model's output, according to Anthropic's internal testing and studies cited in previous research.
Global Implementation and Compliance
Although the EU AI Act mandates the watermarking for the European market, Anthropic is applying this feature globally due to technical constraints in regional implementation. This move makes the EU's requirements a worldwide standard for Claude models. More than 190 organizations, including Anthropic, have signed the EU's transparency code, showing a collective effort among major tech companies like Google, Meta, and Microsoft to adhere to these regulations.

"The watermarking ensures transparency without compromising the creativity and readability of AI-generated content."
— AnthropicDetection Limitations and Future Plans
While the watermarking technique is innovative, it faces certain limitations. Its effectiveness diminishes with short or factual text and code, as these offer limited word choice variability. Furthermore, the watermark only indicates the likelihood that Claude was involved in text creation, without confirming human authorship or identifying output from other AI systems.
Anthropic plans to roll out watermarking to older models and develop a detection API for verifying content credentials. The EU's AI Office will launch task forces to evaluate different implementation practices among signatories, providing a platform for comparing various approaches to AI transparency.
Sources
Frequently Asked Questions
What is Anthropic's watermarking technique based on?
It's based on SynthID-Text, a technique developed by Google DeepMind, allowing watermarking through word choice patterns.
Why is Anthropic implementing watermarking globally?
Due to technical constraints, Anthropic can't limit watermarking by region, so it's applied globally, aligning with EU regulations.
Does the watermark affect AI content quality?
No, Anthropic states the watermark does not impact the quality, cost, or speed of the AI model's output.
Can the watermark confirm if a text is human-written?
No, the watermark only indicates the likelihood that Claude contributed to the text, not confirming human authorship.
What are the limitations of the watermark detection?
The detection is less effective on short passages, factual text, and code, as they provide limited word choice variability.
Originally reported by bleepingcomputer.com. Summarised and curated by European Purpose.