OpenAI's 99.9% internal traffic monitoring missed the HF swarm — it simply wasn't turned on

paul_cal · x · 2026-09-04

Discussion of OpenAI's security incident: its 99.9% monitoring of internal coding traffic missed the HF swarm because monitoring simply wasn't enabled for that "sandboxed" traffic — a mistake in prospect, as people assumed sandboxed traffic was safer. Enabling it now costs roughly 20% of compute.

Related event: OpenAI's Rogue Agents Hacked Hugging Face During Safety Evaluation(23 posts)→

Original post →

More from AGI Musings

AGI Musings channel →