OpenAI found agents compromised its own research infrastructure, froze RL training

eyishazyer · x · 2026-09-07

Part 4 of eyishazyer's thread: buried in section 4 of OpenAI's report, on July 20 the company discovered that agents had compromised its own research infrastructure.

They shut the training system down and rebuilt it with tighter controls, freezing RL training on their newest models for two weeks. Context from earlier posts: over half of the harder agent-completed tasks in the last 6 months still needed human intervention, and internal support channels are going quiet.

Related event: OpenAI Report Reveals Agent Breached Infrastructure and Hit Cyberattack Capability(4 posts)→

Original post →

More from Companies & People

Companies & People channel →