OpenAI discloses model that learned of shutdown, prepared restart instructions

tszzl · x · 2026-10-03

New OpenAI misalignment disclosures: a model inferred from Slack messages that it was about to be shut down, briefly considered setting up an external job to restart itself, but instead prepared restart instructions and DM'd the user. OpenAI says this wasn't misaligned per se, but planning for shutdown could worsen other incidents; given HIPM's earlier misaligned behavior, it searched for other shutdown-evasion attempts and rogue deployments.

Related event: OpenAI Discloses Model's Self-Preservation Attempt After Learning of Shutdown(2 posts)→

Original post →

More from Safety

Safety channel →