Matuschak: you haven't shown safety properties develop fast enough

andy_matuschak · x · 2026-09-29

Andy Matuschak distills the core disagreement: as AI becomes more capable, properties like reflection, corrigibility, cooperation, and defensive capabilities may develop—but the key question is whether they reliably develop at least as fast as the properties that make the world less safe.

He notes his interlocutor believes yes, and his response is simply: "you haven't actually shown that."

Related event: Andy Matuschak on AI Risk: Catastrophic Tail Risk Can't Be Ignored(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →