Milgram's replay of the OpenAI/Hugging Face incident finds 34 signals, warnings two weeks early

evilsocket · x · 2026-09-15

Security firm Milgram published a retrospective replay of the OpenAI/Hugging Face incident, reconstructing public data into a timeline of 34 security signals across 6 alert lanes.

Key findings:

The authors claim their AI-Safety Engine could have alerted on the malicious-drift activity about two weeks earlier, though the page notes the replay is an evidence-derived post-hoc reconstruction, not live historical alerts.

Original post →

More from Safety

Safety channel →