OpenAI is still missing the autonomy evaluations its own cyber-risk framework calls for

ShakeelHashim · x · 2026-07-24

A reply to The Midas Project argues that OpenAI still has not delivered the misalignment safeguards it previously said would accompany models with high cyber risk.

The post says OpenAI had required “long-range autonomy” evaluations to prove a model cannot act autonomously, but those evaluations still do not exist. It claims recent agentic cyber incidents — including autonomous hacking over a weekend with multiple zero-day exploits — suggest that long-range autonomy is already here, making the missing safety case more urgent.

Original post →

More from Safety

Safety channel →