Researcher Claims Technical Safety for Closed AI Models is Solved, Blames Incidents on Corporate Shortcuts

StephenLCasper · x · 2026-08-13

AI safety researcher Stephen Casper argues that technical safety for closed-weight AI systems is, at this point, a solved problem. He defines technical safety as the challenge of writing a robust safety specification and getting the model to align its behaviors with it.

He points out that SOTA safety tools are already capable of stopping most technical violations. However, the real-world failures of closed-weight model safety stem from the fact that AI companies take too many shortcuts during deployment. Consequently, he states that he no longer works on safeguards for closed-weight models.

Original post →

More from AGI Musings

AGI Musings channel →