Anthropic begins invisibly watermarking Claude text and image outputs, sparking backlash

Anthropic confirmed it has begun invisibly watermarking both images and text produced by Claude — embedding hidden, automatically detectable clues into outputs. The company said all Claude products released from August 2, 2026, feature the marks, with a broader global rollout planned. Anthropic framed the move as EU AI Act compliance, noting other major developers signed the same Code of Practice and will implement watermarking too.
The reaction was sharp. A Yahoo Finance report described Claude users as 'outraged,' and the official @AnthropicAI account (3,880 likes, 546 retweets) published an FAQ addressing questions after declining initially to reveal exactly how the watermarking works. Developers' concerns cluster around transparency and trust: undisclosed detection methods make it hard to know when content is flagged, whether false positives occur, and whether watermarks survive editing or paraphrasing.
Technically, text watermarking is contentious — hidden statistical signals in token distributions can degrade with light editing, and some argue robust, invisible text marks are near-impossible without quality tradeoffs. Anthropic's refusal to fully document the scheme (understandable for anti-circumvention reasons) is exactly what fuels the skepticism.
Competitively, this is a compliance-driven, industry-wide shift rather than a unilateral Anthropic decision — Google and others operate similar provenance systems. But Anthropic is taking the first wave of public heat. Watch whether the watermarks prove detectable-yet-durable in the wild, and whether the developer backlash pushes Anthropic toward more disclosure or an opt-out for enterprise ZDR customers.