How Anthropic plans to watermark Claude's AI-generated text
Anthropic, a major AI provider, plans to watermark its AI-generated text to comply with the EU's Code of Practice. The watermark will be invisible and won't affect the quality or content of the text.
Intelligence analysis by Llama

Anthropic is implementing watermarking in its AI model Claude to comply with the EU's Code of Practice. The watermark will be invisible and won't affect the quality or content of the text.
Imagine you're writing a story, and you have to choose between two words to use next. A watermark is like a secret code that you add to the story, but it's not visible to the reader. It's like a hidden message that only a special key can detect. This way, people can know if a story was written by a computer or a human.
Analysis
Anthropic's plan to watermark its AI-generated text is a significant development in the field of artificial intelligence. The company has confirmed that it will implement watermarking across its AI model Claude to comply with the EU's Code of Practice. The watermark will be invisible and won't affect the quality or content of the text, including creativity and readability. According to Anthropic, the watermark will be applied globally at launch because the company doesn't yet have a durable way to scope it by region. The company has also confirmed that future Claude models will generate watermarked text. Models launched before August 2, 2026, are covered by the EU's transition period, and Anthropic is working to add watermarking to those models over the coming months. The implementation of watermarking is based on Google DeepMind's SynthID-Text approach, which works during generation by changing the source of randomness used when making some of the choices. This approach leaves a statistical pattern in the generated text that can be detected by a detector with Anthropic's key. The company has confirmed that internal testing found no impact on creativity, readability, or the content of Claude's responses. Additionally, Anthropic has noted that watermarking requires no additional tokens and has a negligible impact on generation speed. There are certain exceptions to watermarking, including factual statements where only one answer is correct and code, where replacing one term with another could break the output. In these cases, the watermark is not applied. The implementation of watermarking in AI-generated text is a significant development in the field of artificial intelligence, and it will make it easier to identify AI-generated content and ensure transparency in AI-driven processes.
Key points
- Anthropic plans to watermark its AI-generated text to comply with the EU's Code of Practice.
- The watermark will be invisible and won't affect the quality or content of the text.
- The implementation of watermarking is based on Google DeepMind's SynthID-Text approach.
- There are certain exceptions to watermarking, including factual statements and code.
The implementation of watermarking in AI-generated text could lead to increased transparency and trust in AI-driven processes. It could also enable the development of more sophisticated AI models that can generate high-quality content while maintaining the ability to detect their origin.
The implementation of watermarking in AI-generated text could also lead to concerns about the potential misuse of this technology. For example, if the watermark is not robust enough, it could be easily removed or manipulated, which could compromise the integrity of the AI-generated content.

