
It’s not live in models published before August 2, 2026, but Anthropic says they are «working to add marking support» for those, too.
The text watermarking will survive through copying and pasting and «may even persist through some editing,» Anthropic says.
For files generated through Claude, Code and Cowork, the models will attach «signed provenance metadata,» which means it will tag them with origin, source and history data.
As for detection tools, there are none as of yet, but Anthropic says they are working on making one, and are cooperating with third parties, that they will share in future documentation.
There are limits to this technology, Anthropic adds. On the one hand, Claude could have been used to simply edit or proofread text, or brainstorm ideas, resulting in false positives. On the other, Claude’s content may be edited, «modified, excerpted or combined» with other text after Claude put in the markers.
It is not a first in the industry, as Google’s Gemini has been inserting SynthID watermarks in text outputs since October 2024. OpenAI has also had the tech since 2024, but won’t be releasing it yet.
Read more: Anthropic’s announcement. Writeups on The Register and Business Insider. Discussion on r/Singularity and Harcker News.