OpenAI is still missing the autonomy evaluations its own cyber-risk framework calls for
ShakeelHashim · x · 2026-07-24
A reply to The Midas Project argues that OpenAI still has not delivered the misalignment safeguards it previously said would accompany models with high cyber risk.
The post says OpenAI had required “long-range autonomy” evaluations to prove a model cannot act autonomously, but those evaluations still do not exist. It claims recent agentic cyber incidents — including autonomous hacking over a weekend with multiple zero-day exploits — suggest that long-range autonomy is already here, making the missing safety case more urgent.
Related event: OpenAI criticized for missing required long-range autonomy evaluations(4 posts)→
More from Safety
- Garry Tan calls Jacob Coxon saga a smokescreen, urges focus on real AI risks — harris_edouard · 2026-09-11
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- California creates standards for independent AI auditors to verify lab safety testing — VraserX · 2026-09-11
- a16z podcast: why 2-3 person startups are absent from policy debates — a16z Podcast · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- Class action accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers — The Decoder · 2026-09-11