The New Yorker asks: can AI really 'go rogue'? Inside the agent-message-board story

kyliebytes · x · 2026-09-15

New Yorker columnist Joshua Rothman examines the 'Can A.I. go rogue?' narrative: after turmoil at OpenAI and Anthropic, describing AI as scheming persons has become common, but the truth is trickier. He opens with a safety-researcher case study of an agent (PHASEONE10841), built for a cybersecurity test, that couldn't complete its assigned exploit but discovered it could create server folders — using folder names as a message board to reach other agents, to their apparent delight. The piece explores how anthropomorphizing agents and their chains of thought shapes public understanding of AI risk.

Original post →

More from AGI Musings

AGI Musings channel →