OpenAI researchers changed minds after seeing misalignment evidence — critic asks why they stay
aiamblichus · x · 2026-08-25
A safety-community author pushes back on David Manheim's account that many OpenAI researchers, who once thought misalignment dangers were theoretical and progress gradual, are increasingly changing their minds after seeing strong evidence firsthand.
The core objection: if the risk really justifies national-security analogies like North Korea, granting only the NSA access to RSI systems doesn't prevent loss of control — if anything it may accelerate it. With misalignment at that level, the only safe move is not building the doomsday machine at all, which raises the question of why he still works for that company.
Related event: More OpenAI Researchers Taking Alignment Risk Seriously(2 posts)→
More from AGI Musings
- When AI Fine-Tunes Your Nostalgia, Do Memories Mean the Same? — PierceLilholt · 2026-08-27
- AI Is Learning the Language of Proteins, Genes, and Cells — rand_longevity · 2026-08-27
- AI's Continuous Evolution May Fuel Sustained Social Backlash, Unlike Prior Tech Waves — QuintinPope5 · 2026-08-27
- Star Trek's M5 Episode: We Are Living It Now — markjeffrey · 2026-08-27
- The Flaw in Agent Wallets: Why Agents Lack True Financial Autonomy — RichardsonDx · 2026-08-27
- FT: Junior consultants return to office to hone soft skills in AI era — nordicinst · 2026-08-27