mark_k: AI alignment is meaningless unless it means doing exactly what the user intends
mark_k · x · 2026-09-14
markk argues that AI alignment is "bullshit" unless it is defined as doing exactly what the user intends — a pointed take that reframes alignment around instruction-following rather than abstract value alignment, likely to spark debate.
More from AGI Musings
- Rob LeClerc: quoting p(doom) without conditional probabilities reveals shallow thinking — robleclerc · 2026-09-14
- Proposal: an 'aiXiv' venue for AI-written, verified research with no human author — tdietterich · 2026-09-14
- AI-written papers flood arXiv: Dietterich proposes oral exams and an aiXiv venue for AI research — tdietterich · 2026-09-14
- AI Regulation Needs an Asilomar-NPT Playbook, Not a 'China Will Win' Test — krishnan · 2026-09-14
- Math training makes you a stronger LLM user, even beyond math topics — josh_wills · 2026-09-14
- Ethan Mollick: People who never cared about AI are now freaking out about it killing everyone — emollick · 2026-09-14