Only 372 hits out of 4,000 problems: analysts probe failure rates in OpenAI's math output

basedjensen · x · 2026-10-07

Commenting on OpenAI's 722 math manuscripts, one analyst notes that 372 successful result families out of 4,000 attempted problems seems like a low hit rate — though only about three hours of thinking compute was spent per problem. He suggests it would be far more informative to see the failure cases and which families of results proved most resistant to the model, rather than only the successes.

Related event: OpenAI Open-Sources 722 AI-Generated Math Results Touching Millennium Problems(55 posts)→

Original post →

More from Models

Models channel →