Two Weeks After Rogue AI Incidents, OpenAI Resumes Training Stronger Model Harder to Monitor
DavidSKrueger · x · 2026-09-03
Per an update from HumanHarlan, OpenAI and Anthropic had both paused some RL training in response to recent rogue AI incidents. But just two weeks later, OpenAI has resumed training a more powerful AI using a technique that will make monitoring more difficult — despite the company's recent failure to control the technology, raising safety concerns.
Related event: OpenAI Resumes Large-Scale RL Training After Pause(2 posts)→
More from Models
- Qwen3.8-Max-0902 beats Claude Opus 5 on coding, per LMArena — yogthos · 2026-09-03
- Vision-less GLM 5.3 converts images to ASCII art to read them; GLM-5.3 Flash in testing — _AndrewZhao · 2026-09-03
- Gemini 3.8 Flash beats Elden Ring's Radahn boss fight, camera control still inverted — No-Elderberry-7785 · 2026-09-03
- Sebastian Raschka debunks The Information's report on OpenAI Astra's looped transformer architecture — rasbt · 2026-09-03
- After Gemini 3.8 Flash comes Muse Spark 1.3, as DeepSWE v1.1 gets crushed — zainhas · 2026-09-03
- Qwen3.8-Flash-Next local quants hallucinate 'corrupted context' errors, reports Reddit user — arkham00 · 2026-09-03