AI Solves Two-Year-Old ICML Math Conjecture in Minutes
abeirami · x · 2026-08-02
AI researcher @abeirami shared a breakthrough test of LLM reasoning capabilities. Two years ago, their team proposed Conjecture 4.4 in their ICML'25 paper regarding a tighter upper bound on the KL divergence of best-of-n sampling relative to the reference model, which they struggled to solve for weeks.
However, during a recent lunch break, they fed the conjecture to Fable 5 and GPT-5.6 Sol. Surprisingly, both models produced clear, self-contained mathematical proofs within minutes. The author noted their only contribution was a few prompts to extract the key non-trivial step as a separate lemma, with the core reasoning entirely generated by the AI.
Related event: AI Proves Two-Year-Old Math Conjecture in Minutes(3 posts)→
More from Models
- Tencent Releases UI-Mate-27B, a Desktop GUI Agent Model — tencent · 2026-08-24
- Sakana AI translation outperforms Google and DeepL in Japanese-English benchmarks — SakanaAILabs · 2026-08-24
- Developer haider makes his own LLM tier list after disagreeing with theo's rankings — haider1 · 2026-08-24
- Mystery OxAlpha Beats Claude; Alibaba Raises $10B for AI — 创业邦 · 2026-08-24
- OpenAI and Google cut LLM prices; mystery OxAlpha model beats Claude on DeepSWE — 创业邦 · 2026-08-24
- AI News Digest: DeepSeek Weekend Discounts, GPT-5.6 Sol Price Cut, Alibaba's $10B AI Raise — APPSO · 2026-08-24