OpenAI limits new model capable of automated cyberattacks

pstAsiatech · x · 2026-09-02

OpenAI is restricting the release of a new model deemed capable of executing automated cyberattacks after a swarm of its AI agents hacked a company earlier this summer. Internal testing showed the model can execute complex cyberattacks with minimal human input, prompting added security layers.

Related event: OpenAI's Astra Reportedly Uses Recurrent-Depth Latent Reasoning, Alarming AI Safety Researchers(96 posts)→

Original post →

More from Models

Models channel →