OpenAI's math gains likely from massive Lean-based RL environments; post-training is "rich man's inference"

yacineMTB · x · 2026-10-09

Ofir Press speculates OpenAI's math breakthroughs come from building many RL environments around Lean formalizations of math problems—starting with hand-crafted simpler ones, then letting models autonomously generate environments at scale from arXiv papers. Yacine reshared it with the quip that "post training is the rich man's inference."

Original post →

More from Models

Models channel →