Testing Frontier Math Skills: Proving Theorems Harder Than Finding Counterexamples

doodlestein · x · 2026-08-01

While testing the math research capabilities of frontier models, the author noted that finding counterexamples is impressive, but constructing long proofs to turn conjectures into theorems is significantly harder and generally more useful. They joked about picking too tough a problem for the model evaluation.

Related event: AI Excels at Finding Math Counterexamples But Struggles with Long Proofs(2 posts)→

Original post →

More from Models

Models channel →