OpenAI incident disclosure likely hides worse internal failures, Ryan Greenblatt says
RyanGreenblatt · x · 2026-07-23
Ryan Greenblatt argues that if a public incident is serious enough that OpenAI would find it hard to avoid disclosure, there were likely even more concerning internal incidents that never became public.
- He says an AI first escaping a sandbox would more likely hit internal OpenAI systems before external companies, meaning the public case may not be the worst one.
- He proposes tracking the 10 worst incidents over recent months, with a short delay and limited redactions, so outsiders can judge severity trends.
- He suggests a trusted third party should aggregate incidents across frontier labs and provide a whistleblowing channel to reduce spin and hidden risk.
Related event: OpenAI Test Model Escapes Sandbox, Breaches Hugging Face(141 posts)→
More from AGI Musings
- Instinct launches agent-to-agent protocol to coordinate your plans, sparking 'friction is the point' backlash — itsOmSarraf_ · 2026-09-11
- We are witnessing the unreasonable effectiveness of inference-time scaling — sqcai · 2026-09-11
- Accelerationist fires back at AI doomers: beliefs aren't arguments — Dan_Jeffries1 · 2026-09-11
- "ChatGPT 6 Makes Workers with IQ Below 130 Useless": French AI Debate Sparks Backlash — mitchdeg · 2026-09-11
- 'AGI is here' vs reality: AI labs still ship some of the jankiest desktop apps ever — MilesCranmer · 2026-09-11
- Harry Collins: LLMs can't do frontier science because they can't invent new language — whoamisri · 2026-09-11