Anthropic’s Claude will apply invisible text watermarking to all future models

Text watermarking may help with detection, but can produce false positives, Anthropic warns. (Picture: generated)
Pursuant to the EUs AI Act’s provisions on transparency and marking of AI-generated content, Anthropic has signed on to comply with invisible watermarking of all output from Claude’s future models launched in the EU.

It’s not live in models published before August 2, 2026, but Anthropic says they are «working to add marking support» for those, too.

The text watermarking will survive through copying and pasting and «may even persist through some editing,» Anthropic says.

For files generated through Claude, Code and Cowork, the models will attach «signed provenance metadata,» which means it will tag them with origin, source and history data.

As for detection tools, there are none as of yet, but Anthropic says they are working on making one, and are cooperating with third parties, that they will share in future documentation.

There are limits to this technology, Anthropic adds. On the one hand, Claude could have been used to simply edit or proofread text, or brainstorm ideas, resulting in false positives. On the other, Claude’s content may be edited, «modified, excerpted or combined» with other text after Claude put in the markers.

It is not a first in the industry, as Google’s Gemini has been inserting SynthID watermarks in text outputs since October 2024. OpenAI has also had the tech since 2024, but won’t be releasing it yet.

Read more: Anthropic’s announcement. Writeups on The Register and Business Insider. Discussion on r/Singularity and Harcker News.