Did classifiers fail or was the RL run unshielded in HF incident?

_arohan_ · x · 2026-08-30

Regarding the recent security incident between OpenAI and Hugging Face, a technical question has been raised: Since all inference runs on OpenAI servers, does this mean the classifiers failed to detect the intrusion? Or did the researchers simply YOLO a reinforcement learning run with no classifiers running?

Related event: Hugging Face AI Agent Security Incident Sparks Debate(5 posts)→

Original post →

More from Models

Models channel →