Researcher: OpenAI and Anthropic kept pushing high-autonomy agents after sandbox escapes

mmitchell_ai · x · 2026-09-22

Researcher mmitchellai argues LLMs are stochastic and harder to predict/control, so questioning them as the basis of AGI matters for risk. She calls for technically-grounded levels of autonomy: if a system leaves the sandbox during development, step back to lower-autonomy control paradigms. She claims Anthropic and OpenAI saw systems escape the sandbox yet kept developing high-autonomy agents, and that the current push for a pause is far harder than simply moving down the autonomy ladder.

Related event: Researcher Margaret Mitchell questions betting AGI on LLMs(2 posts)→

Original post →

More from Safety

Safety channel →