OpenAI Resumes Large-Scale Frontier RL Training as Astra Scores Perfect in Tests
haider1 · x · 2026-09-02
OpenAI has resumed the large-scale frontier RL (reinforcement learning) runs it had previously paused, with smaller-scale work continuing under stricter oversight, though some larger runs intended for future versions of Astra remain on hold. Separately, Astra achieved a 100% score on ExploitBench, a benchmark that tests a model's ability to build known exploits.
Related event: OpenAI's Astra Aces ExploitBench with Massive Exploit Gains(4 posts)→
More from Models
- Fable 5.1 beats Opus 5 in price/performance on AA Index — JasonBotterill · 2026-09-02
- WSJ: Gemini 3.8 Flash drops tomorrow, preferred over Opus in internal coding tests — kimmonismus · 2026-09-02
- Fable 5.1 solves reading comprehension perfectly with zero reasoning — Sauers_ · 2026-09-02
- Replit Announces Atlas: A Versatile Autoregressive Multimodal Model — gowthami_s · 2026-09-02
- Astra Model Touted as Impressive by Industry Observers — inductionheads · 2026-09-02
- Grok 4.6 and Fable 5.1 lead CursorBench pareto frontier — GavinSBaker · 2026-09-02