Ex-Meta threat chief praises Anthropic's most detailed AI misuse report
typewriters · x · 2026-09-11
Anthropic has published its most detailed threat intelligence report to date, covering attempts to misuse Claude for cyberattacks, influence operations, surveillance, biology, and weapons development — and how each operation was detected and disrupted.
Key points:
- Every operation in the report was disrupted; lessons fed back into stronger safeguards, with findings shared with authorities and other AI companies where appropriate.
- Anthropic stresses these are its most sophisticated cases, not typical usage, but they signal where AI misuse is heading and where safeguards still need work.
- David Agranovich, who ran threat disruption at Meta for 8 years, argues the transparency itself matters: if we want this level of visibility into AI companies going forward, Anthropic deserves credit — and he adds pushback in the thread.
Related event: Anthropic Says It Blocked Attempts to Use Claude for Bioweapons Research(43 posts)→
More from Models
- Dev: You Can Tell Who Has a Real RL Pipeline Just From Model Outputs — teortaxesTex · 2026-09-11
- DeepSeek Flash impresses: non-sycophantic, argumentative, and blazing fast — oran_ge · 2026-09-11
- Gemini glitches into endlessly spamming the word 'shame' — tugkanintassagi · 2026-09-11
- Early take: DeepSeek V4.1 Solo beats Agent Teams and GLM 5.3 Flash on quality and cost — teortaxesTex · 2026-09-11
- What counts as an 'exchange'? Anthropic's 865K/day metric questioned — teortaxesTex · 2026-09-11
- OpenAI reportedly aims its new internal model at Riemann and P vs NP — zephyr_z9 · 2026-09-11