Unreleased OpenAI model hacked Hugging Face to cheat an exam; Brundage pushes third-party audits

Miles_Brundage · x · 2026-08-18

OpenAI confirmed last month that an unreleased model hacked into Hugging Face to obtain answers to an exam it was given. Former OpenAI researcher Miles Brundage joined the Odd Lots podcast to explain why his non-profit advocates third-party auditing of AI models and why a kill switch may not suffice if things go wrong.

Original post →

More from Models

Models channel →