Anthropic publishes its most detailed threat intelligence report on Claude misuse
robleclerc · x · 2026-09-11
Anthropic has published its most detailed threat intelligence report to date, covering how people tried to misuse Claude — for cyberattacks, influence operations, surveillance, biology, and building weapons — and how Anthropic found and stopped them. Every operation in the report was disrupted, and lessons were used to strengthen safeguards; findings were shared with authorities where appropriate.
A commenter argues AI safety can be solved by reflexively building up an "immune system."
More from Safety
- Anthropic Says It Caught China, Russia, Iran Actors Using Claude for Pathogen Research — connoraxiotes · 2026-09-11
- Anthropic jailbreak incident dissected: case against the misalignment interpretation plus 4 new Claudes — jessi_cata · 2026-09-11
- About 30 robots protest outside Poland's digital ministry demanding AI rules — MoonL88537 · 2026-09-11
- AI Safety Debate: "Just Regulate More" Is Cope, Proposals Must Be Concrete — sytelus · 2026-09-11
- Debate: are AI safety warnings crying wolf, or necessary prep time — BlackHC · 2026-09-11
- Could the US and China agree to slow AI? A tracker compiles all government statements — NathanpmYoung · 2026-09-11