Anthropic details watermarking mechanism for Claude-generated text to comply with EU AI Act
The company explains how its SynthID-Text-based watermarks will be embedded in Claude’s outputs, how detectable they are after editing, and how they apply to code.
1 source · cross-referenced
- Anthropic will watermark text generated by its Claude models using the SynthID-Text approach developed by Google DeepMind in 2024.
- Watermarks are undetectable to readers but detectable by parties with a decoding key; they do not degrade output quality.
- Light editing likely won’t remove watermarks, while a full rewrite can; watermarks are less detectable in heavily edited or human-authored text.
- Code watermarks are minimal and limited to comments or arbitrary choices, with negligible impact on functional code.
- Anthropic plans to release a watermark detection API and says other major developers will implement similar watermarks under the EU AI Act’s Transparency Code.
Anthropic published a blog post detailing how it will watermark text generated by its Claude models to comply with the EU AI Act’s Transparency Code, which mandates systems to identify AI-generated content. The company states that watermarks are embedded through patterns in responses that are "undetectable to the reader, but detectable to anyone with a key that encodes it." Anthropic emphasizes that watermarking does not affect output quality and that watermarked responses are indistinguishable from unwatermarked ones to human readers.
The company says it will use the SynthID-Text approach developed by Google DeepMind in 2024, and plans to release a watermark detection API. Anthropic distinguishes its watermarking from AI detection tools like those offered by Pangram, which look for stylistic "tells" in writing rather than checking for a cryptographic watermark.
Anthropic addresses concerns about watermark removal through editing. It says light editing likely won’t remove the watermark, while a complete rewrite that replaces every word can. In the latter case, the company argues, the resulting text may no longer qualify as AI-generated. For text proofread or lightly edited by Claude, detectability depends on how much of the content remains original; if most words are human-authored, there is little for the watermark to attach to.
The company notes that code will generally carry fewer detectable watermarks because functional constraints limit the model’s freedom to choose among equivalent options. Watermarks may appear in comments or areas where arbitrary word choices exist, but their impact on the actual code is described as negligible.
Anthropic states that other major model developers have signed the same Code of Practice and will implement their own watermarks, indicating a broader industry move toward standardized content provenance under the EU AI Act.
- Aug 15, 2026 · TechCrunch — AI
French startup Kog targets faster LLM inference on existing GPUs with low-level optimization
Trust72 - Aug 14, 2026 · Hugging Face
Hugging Face finds Chinese labs lead open model releases at frontier scale while U.S. hardware vendors dominate new model uploads
Trust79 - Aug 14, 2026 · TechCrunch — AI
Google adds toggle to remove visible AI watermarks from images, video, and audio
Trust79