Expert Doubts Anthropic Watermark: Stochastic Fingerprint as Unreliable as Hallucinations

gerardsans · x · 2026-08-13

A tweet quotes gerardsans' comment that Anthropic's watermarking is fundamentally flawed: the Claude fingerprint is stochastic, as unreliable as hallucinations, and such methods were abandoned due to false positives. It can only tell if the fingerprint matches the Claude family with ample margin of error. Certification methods like DRM are more reliable, but AI deals with text only on the consumer side. Awaiting Anthropic's details, but the current method is imperfect by design.

Related event: Experts Question AI Watermarks, Advocate DRM(2 posts)→

Original post →

More from Safety

Safety channel →