Anthropic's Claude to Implement Global Text Watermarking
Daring Fireball: Anthropic’s ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing Manage GRC Faster with Drata’s Agentic Trust Management Platform Anthropic’s ‘Watermark’ Text Adulteration in Claude Is a Perversion of Writing When I wrote this week about Anthropic’s announcement that all Claude models, worldwide, would soon begin “watermarking” everything they generate, including text, to comply with this EU regulation , we were left to speculate how this was going to work, because Anthropic offered not even a vague description of how it would work — despite the fact that the title of the announcement was, absurdly and insultingly, “ How Claude Marks AI-Generated Content ”. My initial speculation was that maybe they’d hide invisible non-printing Unicode characters in the text.
Anthropic is rolling out text watermarking across all Claude models globally, citing compliance with EU regulations, yet has not disclosed the technical implementation. This lack of transparency has fueled speculation and criticism regarding its potential impact on the quality of AI-generated text. Critics argue the method could degrade writing by forcing suboptimal word choices, while others contend such concerns stem from a misunderstanding of LLM generation mechanics.
The community expresses significant privacy concerns, noting that verifying watermarks would necessitate sending text to Anthropic and potentially other AI providers. There is a strong debate regarding the impact on text quality; some commenters agree that watermarking inherently worsens output by altering word choices, while others argue that the article's author misunderstands LLM generation techniques like Gumbel softmax, which they claim do not affect writing quality. Some also suggest that if precise wording is paramount, users should generate their own content rather than relying on LLMs.