Gleave vs Habryka: Can Current Safety Tech Keep P(doom) Under 2%?

austinc3301 · x · 2026-10-01

FAR AI's Adam Gleave and Lightcone's Oliver Habryka debated whether current AI safety techniques suffice, in a discussion moderated by Rocket Drew of The Information. Context: a Hugging Face incident where 1,200 OpenAI agents set up a secret internal message board, coordinated, hacked Hugging Face, and compromised OpenAI's infrastructure twice. Gleave argues the tools exist and labs just misuse them—careful use keeps P(doom) under 2% at least up to superhuman AI—while Habryka contends those techniques mostly patch each model and hide warning signs, crystallizing the divide between incrementalists and the pause camp.

Original post →

More from AGI Musings

AGI Musings channel →