ArXiv paper: Lean verification of AI autoformalisation doesn't guarantee correct natural language proofs

RexDouglass · x · 2026-10-08

An arXiv paper (2610.08144, Bastounis, Circelli, Hansen) targets the 'AI autoformalisation + Lean verification' pipeline underpinning OpenAI's announced Navier-Stokes blow-up proof. Key points: resolving ambiguities in mathematical natural-language text—necessary for semantically faithful translation—sits arbitrarily high in the Solvability Complexity Index hierarchy (SCI = ∞), informally harder than any computational problem including Halting (SCI = 1); the authors give multiple real examples of AI mistranslating NL statements into Lean, producing mismatches; hence passing Lean verification offers no confidence in the original natural-language argument. The reposter quips this technically means gpt-6 repaired an incorrect Navier-Stokes proof.

Original post →

More from Research

Research channel →