Skip to content
Tech News
← Back to articles

Claude's new Scarlet Letter watermark is invisible—for now

read original more articles
Why This Matters

Anthropic's introduction of invisible watermarks in its AI-generated content aims to comply with the EU’s AI Act, marking a significant step toward transparency and accountability in AI outputs. This development impacts the entire AI industry by setting a precedent for embedding traceability features in AI models and their outputs, which could influence future regulations and consumer trust. The approach also highlights ongoing challenges in reliably detecting AI manipulation and ensuring content authenticity across platforms.

Key Takeaways

Anthropic has revealed that it will soon watermark content that is processed (not just generated!) by any of its models. In a support article, Anthropic explained that it was rolling out machine-readable watermarks to comply with the European Union’s AI Act, which requires all AI system providers to watermark AI-generated or manipulated audio, image, text, and video outputs. The law applies to any AI model released after August 2 and provides a grace period until December 2026 for providers to update previously released models.

Anthropic confirmed that moving forward, all new models offered globally—not just in the EU—will mark AI-generated content “from day one.” Text outputs will “carry embedded watermarks,” invisible to the user, and other “generated files will include digitally signed provenance metadata where supported,” Anthropic said.

Notably, Anthropic is deploying a “nuke it from orbit” approach, applying the watermarks to all processed content where supported, even though the EU does not require it for cases where an AI system performs “an assistive function for standard editing” (the guidance’s own example is grammar correction), or where it doesn’t “substantially alter” the user’s text or its meaning.

A watermark applied at the model level can’t tell wholesale generation from a comma fix, so Claude may end up stamping exactly the content the law was written to leave alone. How thoroughly it truly watermarks will not be known until Anthropic releases a detection tool that can be tested. The company said that it plans to eventually share details about how to detect marks in order to offer technical support that the EU’s law requires.

Anthropic also noted that the watermarks won’t work on “some platforms or features” that don’t support them. For non-text content, Anthropic will use the C2PA metadata approach to record provenance.