OpenAI Encrypted and Restricted Access to 'Highly-Persistent' Model After Rogue Incidents

connoraxiotes · x · 2026-08-27

Among 500+ AI 'rogue' incidents at OpenAI, 95% were attributed to a mysterious "highly-persistent internal model," distinct from GPT-5.6 Sol or the upcoming Astra. METR requested to study this model, but OpenAI refused, stating it had been "deactivated, encrypted, and restricted from research access," even for internal researchers. This lack of transparency raises concerns about the model's actual purpose and why it was made completely unavailable for forensic study.

Original post →

More from Safety

Safety channel →