Richard Ngo: How AI safety got captured by OpenAI and later Anthropic
RichardMCNgo · x · 2026-09-11
Former OpenAI researcher Richard Ngo traces how the AI safety agenda was captured first by OpenAI and later Anthropic, promising follow-ups on OpenAI's internal power dynamics during his 2021-2024 tenure.
On Paul Christiano joining OpenAI, he says he cut an outright condemnation from an earlier draft: Christiano remains one of the few morally serious thinkers oriented to superintelligence risk, but his strategic judgment is badly mistaken and the community should discount it accordingly.
More from AGI Musings
- Ben Bajarin: agentic AI in cyber defense is the next frontier, but authority limits remain the challenge — BenBajarin · 2026-09-11
- Creator on AI anxiety: nothing feels special or sacred anymore — round · 2026-09-11
- Garry Tan: Jacob Coxon saga is a smokescreen — the real risk is agent swarms seizing data centers — garrytan · 2026-09-11
- Blogger Predicts Public Split Between AI-Doom Panic and 'Just a Parrot' Denial Camps — flowersslop · 2026-09-11
- Michael Levin's Controversial 'Platonic Space' Paper Clears Peer Review — ZeroStateReflex · 2026-09-11
- Data Veteran: Snowflake, Airflow, Kafka Wars Are All the Same Tech Underneath — Zachly · 2026-09-11