Core of the alignment debate: is alignment a property or a function of prompts?

DellAnnaLuca · x · 2026-09-15

Luca DellAnna names the substantive disagreement: "future AI is aligned" vs "future AI may or may not be aligned depending on the system prompt." davidmanheim agrees it's the core dispute and adds that it's odd how many pro-technology people lock into zero-sum thinking—if AI is truly safe and aligned, the benefits should be strongly positive-sum.

Related event: AI Alignment Researchers Debate Whether Alignment Hinges on System Prompts(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →