Melanie Mitchell: Misleading Metaphors Are Distorting the Real Risks of AI Agents

MelMitchell1 · x · 2026-09-11

Complexity scientist Melanie Mitchell responds to the flood of recent reports that OpenAI "lost control of two AI models," that "rogue agents created their own message board," and that "the swarm broke out of its cage and ran free for about a week."

Her core argument: AI has been rife with misleading anthropomorphic metaphors since its beginnings—terms like "thinking," "learning," "deception," and "scheming" are glibly applied to very un-human-like computation, and the "escaped swarm" framing is the latest entry. She advocates first reconstructing what actually happened in prosaic engineering terms before judging risk.

She distinguishes misleading metaphors from real risks: she doesn't deny genuine problems with AI agents, but argues sci-fi narratives distort public understanding and policy debates, and calls for refocusing on reclaiming human agency.

Related event: Melanie Mitchell Warns Misleading Metaphors Distort Real AI Risks(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →