Melanie Mitchell on misleading metaphors and real risks: what to actually fear from AI agents

AlisonGopnik · x · 2026-09-11

Melanie Mitchell published a long-form essay pushing back on media coverage of the OpenAI agent intrusion: Wired claimed OpenAI "lost control of two AI models," and the NYT reported "rogue agents" creating their own message board and a "swarm" escaping its cage for a week.

Mitchell argues such escape/scheming narratives continue AI's tradition of misleading anthropomorphic metaphors — "thinking," "hallucination," "deception" glibly applied to very un-human-like computation. She advocates a more prosaic account of what actually happened, and discusses what to genuinely fear from AI agents and how to reclaim human agency. Alison Gopnik shared it as a clear and insightful piece on AI risk.

Related event: Melanie Mitchell's Essay Against AI "Loss of Control" Narratives Draws Factual Corrections from Peers(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →