Anthropic publishes its most detailed threat intelligence report on Claude misuse

robleclerc · x · 2026-09-11

Anthropic has published its most detailed threat intelligence report to date, covering how people tried to misuse Claude — for cyberattacks, influence operations, surveillance, biology, and building weapons — and how Anthropic found and stopped them. Every operation in the report was disrupted, and lessons were used to strengthen safeguards; findings were shared with authorities where appropriate.

A commenter argues AI safety can be solved by reflexively building up an "immune system."

Related event: Anthropic Releases Its Most Detailed Threat Intelligence Report on Claude Misuse(16 posts)→

Original post →

More from Safety

Safety channel →