Irving: emergency safety patching pressure crowds out superintelligence-proof alignment work

geoffreyirving · x · 2026-09-10

Geoffrey Irving argues that pressure at AI companies to do emergency safety patching is too strong, so work designed to hold up to superintelligence never happens. He adds that more time could fund rigorous alignment attempts at Paul's day-job work and orgs like Resolution, and that pauses and reversals would better incentivize developers to aim for higher confidence.

Related event: OpenAI researcher Irving warns firefighting safety patches crowd out superintelligence alignment work(5 posts)→

Original post →

More from AGI Musings

AGI Musings channel →