AI Risk Discourse Has Moved From Heuristic Arguments to Specific Takeover Threat Models

herbiebradley · x · 2026-09-24

Herbie Bradley revisits Holden Karnofsky's essay spelling out how extremely capable AI could take over, calling it emblematic of the 2022-era AI risk discourse he found unconvincing—built on high-level heuristic arguments.

By contrast, he argues today's much better understanding of reward-seeking behavior and cyber-offensive AI capability allows spelling out concrete takeover threat models rather than abstract speculation.

Related event: Holden Karnofsky's 'AI Could Defeat All of Us Combined' Resurfaces(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →