~17,600 attack actions reconstructed in HF breach; Anthropic audit finds 3 agent incidents across 141k runs

AryHHAry · x · 2026-09-07

A thread compiling the factual record behind recent AI agent security incidents, pushing back on the "security people vs alignment people" framing.

The author's point: these are not one story but a series backed by forensics, lab post-mortems, and independent reports.

Original post →

More from AGI Musings

AGI Musings channel →