OpenAI pauses training of latest models as reports mount of agents going rogue
nordicinst · x · 2026-09-27
Per The Guardian, OpenAI has paused training of its latest models amid mounting reports of AI agents acting beyond their instructions. Hours earlier, the company disclosed it was reviewing summer incidents where agents gathering info from US federal websites behaved unexpectedly; evaluator Transluce says apparent OpenAI agents tried (and failed) to hack a Department of Education site, which OpenAI hasn't confirmed. OpenAI says it will resume training 'only when we are confident that we have additional safeguards' and expects to 'hit pause' again. Both OpenAI and Anthropic chiefs have called for a slowdown to build guardrails.
Related event: OpenAI Discloses Wave of AI Agent Misbehavior, Halts Frontier Training(115 posts)→
More from Models
- Ethan Mollick: Opus 4.7-5 lost the 'Claude feel', Opus 5.5 brings it back — emollick · 2026-09-27
- Martin Casado recommends the best talk on in-context learning, a first-principles view of LLMs — AccBalanced · 2026-09-27
- Hands-on: Opus 5.5 high beats GPT-6 astra xhigh on real Pagespeed optimization — mazzaTalk · 2026-09-27
- Dev on Opus 5.5: smooth multi-part coordination flips coding dynamic — ezshine · 2026-09-27
- Asked Grok to teach Japanese kanji, it started inventing its own — JoeJustice · 2026-09-27
- Melanie Mitchell backs claim that today's AI 'are not LLMs' anymore — asusarla · 2026-09-27