Kaj Sotala argues alignment should favor wise-advisor designs over untested CEV

xuenay · x · 2026-09-09

In a debate with Raemon about alignment approaches, researcher Kaj Sotala lays out his stance: be careful and stick to minimal solutions close to what's known to work. He argues we don't know how to build a superintelligence that would stick to a wise advisor role, but we do roughly know what wise advisors are and should be like—making it a safer baseline than speculative schemes.

Related event: Alignment Researchers Push Back on CEV as Unproven and Risky(4 posts)→

Original post →

More from AGI Musings

AGI Musings channel →