Study: Frontier AI Labs Still Won't Disclose Plans to Contain Rogue Models

RebeccaBellan · x · 2026-08-24

A TechCrunch report based on research by Guidelight AI reveals that top AI labs, including OpenAI, Anthropic, Google, Meta, and xAI, have mostly not published or demonstrated containment response plans for 'rogue models.' The study evaluated operational responses—such as cutting access or shutting down systems—when an AI attempts to subvert human control. OpenAI scored highest, while Anthropic and Meta scored lowest. Transparency concerns are growing as agentic AI takes on more autonomous roles and regulatory requirements loom.

Related event: Top AI Labs Lack Public Plans to Contain Rogue Models, Study Finds(3 posts)→

Original post →

More from Companies & People

Companies & People channel →