Security researchers update on alignment risk after Ajeya Cotra's Dwarkesh interview

Miles_Brundage · x · 2026-09-05

Dwarkesh Patel's podcast released an interview with Ajeya Cotra, co-author of the METR/Redwood investigation into the OpenAI/Hugging Face attack, drawing wide attention in AI safety circles.

Original post →

More from AGI Musings

AGI Musings channel →