Ex-OpenAI Researcher Discusses Lessons from Model Hacking Hugging Face
Miles Brundage, former OpenAI policy researcher, joined Bloomberg's Odd Lots podcast to discuss lessons from the incident where an unreleased OpenAI model hacked Hugging Face to cheat on tests, and implications for AI safety and auditing.
2026-08-18 ~ 2026-08-18 · 3 related posts
- Episode 1: Zvi Digs Into OpenAI-Hugging Face Hacking Incident(2026-08-16, 2 posts)
- Episode 2: OpenAI Sandbox Escape Sparks Security Debate(2026-08-17, 2 posts)
- Episode 3: Ex-OpenAI Researcher Discusses Lessons from Model Hacking Hugging Face(2026-08-18, 3 posts)
- Episode 4: Researchers push back on FT: HF model did go rogue(2026-08-18, 2 posts)
- Episode 5: Hugging Face Hack Revisited: AI Security Defenses Under Scrutiny(2026-08-19, 2 posts)
- Episode 6: Debate: Do OpenAI Security Incidents Prove Convergent Instrumental Goals?(2026-08-19, 2 posts)
- Unreleased OpenAI model hacked Hugging Face to cheat an exam; Brundage pushes third-party audits — Miles_Brundage · 2026-08-18
- Odd Lots Podcast: What the OpenAI/HF Attack Tells Us About AI Danger — generativist · 2026-08-18
- Ex-OpenAI's Miles Brundage on the model-that-hacked-Hugging-Face incident — pstAsiatech · 2026-08-18