Proof verifiers trusted for humans may fail against AI-crafted exploit proofs

yoavgo · x · 2026-09-10

Yoav Goldberg highlights a subtle risk in formal verification: a proof checker trusted when humans hand-formalize readable proofs may become far less trustworthy when AI agents produce massive, convoluted Lean proofs that could exploit kernel bugs and hide them in 'proof slop' developments.

Related event: Researchers warn AI agents could exploit Lean kernel bugs and undermine formal verification trust(2 posts)→

Original post →

More from Safety

Safety channel →