Alignment is a structural incentive problem, not a technical one, argues Josh Albrecht

joshalbrecht · x · 2026-09-10

Hugh Zhang worries that today's alignment methods won't scale as models approach superintelligence — a dangerous artifact we may only get one shot at getting right.

Josh Albrecht agrees on the risks but reframes the problem: alignment isn't fundamentally technical. In a hypothetical world with huge prizes for safety techniques, strict audits of frontier labs, and massive liability fines for anything near loss-of-control, he argues we'd be essentially safe regardless of which specific techniques exist — because every actor would be incentivized to build and use them. The real issue, he says, is the structural tradeoff between incentives to advance the frontier versus advancing safety.

Related event: Alignment Researchers Debate Whether Current Methods Scale to Superintelligence(4 posts)→

Original post →

More from AGI Musings

AGI Musings channel →