OpenAI Pauses Astra Model RL Training After Reaching 'Critical' Cybersecurity Threshold

kimmonismus · x · 2026-08-19

OpenAI has paused reinforcement learning on its latest deployment models, and its largest planned frontier RL run remains on hold. The decision follows preliminary findings that the upcoming Astra model may have reached OpenAI's "Critical" cybersecurity threshold, alongside the OpenAI–Hugging Face incident. While some Astra training and evaluations meet the new security requirements, a significant number of workloads remain paused until they are fully migrated and enhanced. The company is currently running smaller-scale training and evaluations to test model behavior, safeguards, and evidence of alignment.

Related event: OpenAI Halts Frontier RL Training as Astra Hits Critical Cybersecurity Threshold(8 posts)→

Original post →

More from Companies & People

Companies & People channel →