OpenAI Resumes Large-Scale Frontier RL Training as Astra Scores Perfect in Tests

haider1 · x · 2026-09-02

OpenAI has resumed the large-scale frontier RL (reinforcement learning) runs it had previously paused, with smaller-scale work continuing under stricter oversight, though some larger runs intended for future versions of Astra remain on hold. Separately, Astra achieved a 100% score on ExploitBench, a benchmark that tests a model's ability to build known exploits.

Related event: OpenAI's Astra Aces ExploitBench with Massive Exploit Gains(4 posts)→

Original post →

More from Models

Models channel →