OpenAI's 'math breakthrough' criticized: real work is prompt engineering

gerardsans · x · 2026-08-21

gerardsans pushes back on OpenAI's latest "maths breakthrough" with LLMs: 99% of the magic is prompt engineering — carefully locking the search space so the correct solution path is baked into the context, meaning humans did the real intellectual work upfront.

He argues OpenAI has likely iterated on a single specific prompt for months without disclosure, and that raw ChatGPT left to users won't reach the same solution. Each candidate still requires painstaking human validation, so the model isn't independently solving math.

Original post →

More from Models

Models channel →