Anthropic's Will Carlson: AI content isn't always detectable, making policy hard

willcb · x · 2026-10-03

Replying to a discussion on regulating AI-generated content, Anthropic's willcb argues existing defamation frameworks remain useful and platforms already use watermarks for labeling — but we can't assume all AI-generated content is black-box detectable, or even clearly definable, which makes policy hard.

Related event: AI-Generated Content Is Hard to Detect Reliably, Fueling Policy Debate(2 posts)→

Original post →

More from Safety

Safety channel →