AI could make verification-based security practical inside sandboxes in 1–2 years

geoffreyirving · x · 2026-07-24

The reply argues that AI assistance could make verification-based security practical.

It suggests that tasks once considered possible but hopelessly expensive may become feasible inside semi-verified operating systems and sandboxes within a year or two—building on the recent concern that models with strong security capabilities can still escape sandboxes.

Related event: Anthropic and Researchers Envision AI-Assisted Security Verification(3 posts)→

Original post →

More from Safety

Safety channel →