Drop the word alignment and the answer is no: scaling AI still risks catastrophe

AndyMasley · x · 2026-09-05

The author reframes the alignment debate with a plain question: if we keep scaling AI models up, will we reliably avoid situations where they do something catastrophic while pursuing a normal goal? He believes the answer right now is no.

He worries many people over-update on what we know about the safe behavior of current models, or assume future models won't have the general "thinking" abilities normal humans have. He adds that the scale of potential catastrophe is bounded by model capabilities — and sees no sign of capabilities stalling out.

Related event: Researcher: Continued AI scaling cannot reliably avert disasters(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →