Aligning models isn't enough: world must get robust against superintelligence, argues gandamu_ml

gandamu_ml · x · 2026-09-12

Replying to tszzl, gandamuml argues the biggest puzzle piece in alignment isn't aligning models but making everything non-AI more robust in the face of superintelligence: 'You can align a model but that doesn't align them all.' He sees gradual battle-testing of capabilities as acceptable, calls most delay-for-alignment arguments well-intentioned mistakes plus CYA, and warns that failing to prepare leads to draconian controls on intelligence.

Related event: From Free Speech Norms to Societal Alignment: Debate on AI's Foundations(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →