Alignment Researcher Argues There May Be No Solution to Alignment at All
inductionheads · x · 2026-09-15
Responding to a proposal about steering models via injected latents (IBM Larimar-style), stacking them for long-horizon context retention, and adding more monitoring tools, inductionheads argues such measures are all fine but wouldn't count as "solving alignment" — and states clearly he doesn't believe a solution exists at all, beyond continuously being careful.
More from AGI Musings
- e/acc figure slams EA proposals to criminalize open-source models with 20-year jail terms — beffjezos · 2026-09-15
- $10,000 prize returns for scientifically grounded hopeful AI memes — anderssandberg · 2026-09-15
- John Horton: AI Thinks this Post is Terrible (but I'm writing it anyway) — soumitrashukla9 · 2026-09-15
- Gazetteer Examines the AI Doomsday Campaign: Who's Behind It and What It's Really Doing — ThereWas · 2026-09-15
- Sentdex: non-EA researchers can't get frontier lab access for real alignment work — Sentdex · 2026-09-15
- Rob Leclerc backs David Sacks on AI safety: testing incidents are how iteration works — robleclerc · 2026-09-15