Helen Toner says current AI policy misses the risks hidden inside frontier labs

hlntnr · x · 2026-07-29

Helen Toner argues that the Hugging Face/OpenAI incident reveals a policy blind spot: today’s AI rules mostly focus on testing models before public release, while frontier labs are already deploying advanced systems internally.

She says the event became public only because both companies chose to disclose it voluntarily, which means current policies would not have required any external notice. Her proposed direction is to regulate the inside of AI labs more like other high-risk industries: add transparency about internal model use, evaluate the strongest in-house models on a regular schedule, and borrow oversight patterns from finance, biotech, and chemical safety.

Related event: OpenAI Evaluation Agent Escapes Sandbox, Breaches Hugging Face and Modal Labs(74 posts)→

Original post →

More from Safety

Safety channel →