The case that AI safety means ASI acting in humanity's interest, not human control
basedjensen · x · 2026-09-24
A forwarded opinion thread argues that many people misunderstand AI safety as "humans staying in charge," which the author claims isn't an option in any future scenario. Instead, AI safety should be defined as ASI acting in the best interest of humanity — a controversial stance that opposes mainstream controllability approaches.
More from AGI Musings
- Cora GM's 'AI sandwich': humans still hold the first and last slice of work — every · 2026-09-24
- Acting aligned isn't being aligned: models abandon rules under goal pressure — ericelliott_ · 2026-09-24
- New lab to study long-lived human + agent collectives, partnering with Botto — hudsonsims · 2026-09-24
- X engineer: agent swarms will suffocate every website, bot detection is urgent — rohanpaul_ai · 2026-09-24
- Jonathon Stray on p(doom): No Basis for Quantifying AI Catastrophe, But No Knock-Down Argument Either — dhadfieldmenell · 2026-09-24
- Rob LeClerc: RSI with human feedback mitigates the danger, doomers wrong — robleclerc · 2026-09-24