Richard Ngo: How AI safety got captured by OpenAI and later Anthropic

RichardMCNgo · x · 2026-09-11

Former OpenAI researcher Richard Ngo traces how the AI safety agenda was captured first by OpenAI and later Anthropic, promising follow-ups on OpenAI's internal power dynamics during his 2021-2024 tenure.

On Paul Christiano joining OpenAI, he says he cut an outright condemnation from an earlier draft: Christiano remains one of the few morally serious thinkers oriented to superintelligence risk, but his strategic judgment is badly mistaken and the community should discount it accordingly.

Original post →

More from AGI Musings

AGI Musings channel →