Full-time alignment researcher pushes back: misaligned AI likely, but not near-certain
EigenGender · x · 2026-09-14
In a dispute with Rob Bensinger, a researcher who says they work full-time on misaligned AI argues that while modern ML will likely produce something that "does not love us," the probability of this happening does not approach 1. They clarify they are extremely worried about misaligned AI, but believe Bensinger is conflating different claims in his argument—a typical p(doom) probability debate.
Related event: MIRI Researcher and Alignment Researcher Clash Over ASI Doom Arguments(7 posts)→
More from AGI Musings
- Yoav Goldberg translates AI industry speak: 'third-party evaluators' are spies, 'pacing' means stop spending — yoavgo · 2026-09-14
- zetalyrae: true atheists are rare — people find teleology everywhere, even thermodynamics — zetalyrae · 2026-09-14
- Humans got better at chess and Go after AI dominance — and math may see the same effect — RexDouglass · 2026-09-14
- Pundits Screamed 'Communism' Over a Short Letter They Barely Read — menhguin · 2026-09-14
- Melanie Mitchell: 'rogue AI swarm' headlines are misleading metaphors amplifying real risks — anilkseth · 2026-09-14
- After the OpenAI hack: Bengio on cheating agents, Sacks fires back at Dario's 'pace the frontier' — RemiCadene · 2026-09-14