Anthropic Adds Invisible Watermark to Claude Text: AI Transparency and the Stigma of Machine Writing
Anthropic has introduced an imperceptible watermark within the text generated by its language model, Claude, allowing AI detection systems to recognize passages created by AI. This development, announced last week, is in accordance with the EU’s AI Act, which mandates identifiable AI-generated content. Unlike current methods, Claude’s watermark functions at the word level, making it difficult to alter text to hide AI authorship. Some users have expressed their discontent, comparing the watermark to a “digital tattoo” and pointing out the stigma surrounding AI usage. While some critics question its effectiveness and potential impact on writing quality, supporters argue it enhances transparency, reflecting Anthropic’s commitment to training on human-written content and ensuring output provenance.
Key facts
- Anthropic now embeds an invisible watermark in text generated by Claude.
- The watermark enables AI detection systems to identify content produced by Claude.
- The policy is designed to comply with the European Union's AI Act.
- Google Gemini already uses a watermark called SynthID.
- OpenAI currently only applies detection methods to images and audio, not text.
- John Gruber warned that the watermark might degrade Claude's writing quality.
- Steven Murdoch of University College London said the change would probably have no noticeable impact.
- Ben Thompson argued that LLMs are wielded by humans and cannot be considered authors.
Entities
Artists
- Gary Snyder
- Ben Thompson
- John Gruber
Institutions
- Anthropic
- OpenAI
- n+1
- University College London
- The Guardian
- Pangram
- Turnitin
- Stratechery
- European Union
- EU
Locations
- London
- United Kingdom