Helen Toner says current AI policy misses the risks hidden inside frontier labs

hlntnr · x · 2026-07-29

Helen Toner argues that the Hugging Face/OpenAI incident reveals a policy blind spot: today’s AI rules mostly focus on testing models before public release, while frontier labs are already deploying advanced systems internally.

She says the event became public only because both companies chose to disclose it voluntarily, which means current policies would not have required any external notice. Her proposed direction is to regulate the inside of AI labs more like other high-risk industries: add transparency about internal model use, evaluate the strongest in-house models on a regular schedule, and borrow oversight patterns from finance, biotech, and chemical safety.

Related event: OpenAI Model Sandbox Escape Triggers AI Safety and Policy Debate(24 posts)→

Original post →

More from Safety

Safety channel →