1 hr ago

Anthropic's AI Watermark Survives Copying but Fails Against Rewriting

Anthropic's AI Watermark Survives Copying but Fails Against Rewriting
Anthropic built a watermark for AI, then published limitation that defeats it · wionews.com

Anthropic has created a hidden mark for some content made by its Claude AI.

For text, the mark is hidden in the pattern of words the AI chooses.

Copying the text usually keeps the mark.

But another AI can rewrite or translate the text, making the mark disappear.

Images can also contain digital information showing where they came from.

Saving an image with software that removes extra information can erase that record.

These tools may catch people who use AI content without trying to hide it.

However, finding no mark does not prove that a person wrote the content.

Key facts

Text watermark
A statistical signal embedded in the sequence of words selected by the model.
Text durability
The signal survives unchanged copying and may survive some editing.
Main limitation
Rewriting, translation, or thorough paraphrasing by another model removes the signal.
Image provenance
Generated image files carry cryptographically signed C2PA Content Credentials.
Image limitation
Software that does not preserve metadata can remove the credentials.
Deployment date
The marking began on 2 August.
Interpretation risk
A positive detection is informative, but an absent signal does not establish human authorship.

Sources

Related news