Claude to embed imperceptible C2PA watermarks in AI-generated text and images

Anthropic's move makes it one of the first frontier labs to bake provenance signals directly into generated text at scale. The watermark travels with copied text and may persist through some editing, aiming to police 'AI slop' and undetected AI writing. For images (SVG, PNG, JPG), Anthropic attaches signed C2PA provenance metadata, aligning with the emerging cross-industry standard backed by Adobe, Microsoft and others.
The trigger is regulatory: the EU AI Act's Article 50(2) obliges providers to mark AI-generated content in machine-readable form. Anthropic scoped the commitment to models launched on or after August 2, meaning older Claude versions are not retroactively watermarked — a caveat critics note limits near-term coverage. Enforcement therefore ramps only as new models roll out.
Competitively, this raises pressure on OpenAI, Google and Meta to match verifiable provenance, particularly as regulators worldwide watch the EU's implementation. The C2PA image path is well-trodden, but robust, copy-paste-surviving text watermarking is technically harder and more novel — and the most scrutinized part of the announcement.
Skeptics are vocal. Developer Nick Dobos said 'I don't want invisible information in my codebase that I don't control,' and writers who rely on undetected AI-generated prose fear a mark that survives copy-paste and some editing. Others question how robust the text watermark really is against paraphrasing or adversarial stripping. Anthropic frames it as a transparency win; the open question is whether the marks hold up outside controlled conditions.