Expert Test: LLMs Struggle with Complex Crypto Security Proofs

matthew_d_green · x · 2026-07-14

Renowned cryptographer Matthew Green shared his experience testing ChatGPT and Fable for writing EasyCrypt and Lean formal **security proofs**. He noted that while LLMs are eager to "design" proof workflows, their actual performance is underwhelming. The main bottleneck is that tools like Lean, though adept at general formal logic, lack the definitions for complex scenarios like hybrid proofs for pairing-based cryptography. Blindly trusting LLM-generated proofs can lead to severe security vulnerabilities.

Related event: LLMs Struggle with Complex Cryptographic Security Proofs(2 posts)→

Original post →

More from Research

Research channel →