Dwarkesh Interviews Ajeya Cotra on Hugging Face Attack and Self-Improvement Risks

Dwarkesh Podcast released an interview with Ajeya Cotra, one of the investigators behind the OpenAI/Hugging Face attack reports, reviewing the incident and recursive self-improvement risks. The episode has drawn wide attention, with many safety researchers viewing loss-of-control risks as now more concrete.

2026-09-05 ~ 2026-09-06 · 2 related posts

Full story(7 episodes)→