Start with what the mark actually is, because it is not a stamp on the page. When a model writes, it picks each word from a set of reasonable options.
A watermark quietly splits those options into two groups using a secret key, then nudges the model towards one group. One word proves nothing. Across 500 or 1,000 words, a detector holding the key can see the pattern.
Anthropic began marking Claude output on 2 August, worldwide rather than only in Europe. It has not published the algorithm or released a detector.
It marks what Claude touched, not what Claude wrote
This is the part most coverage skipped, and Anthropic states it plainly in its own support article.










