ICML Research: Proof Verification Scaling and Long-Horizon Planning
VectorInst · x · 2026-07-08
Other ICML research highlights this week include: Renjie Liao on scaling mathematical proof verification; Florian Shkurti on stable offline-to-online reinforcement learning transfer; and Igor Gilitschenski on test-time planning, which successfully improved the success rate of long-horizon tasks from near zero to over 90%.
Related event: Research Unravels GRPO Training Collapse; ICML Highlights AI Scaling(2 posts)→
More from Research
- Kimi K3 and Fable 5 now look much closer than the old open-vs-closed gap — FinanceYF5 · 2026-07-21
- uv-scripts/ocr returns to the top of Hugging Face datasets with a JSON model picker — vanstriendaniel · 2026-07-21
- DeepSearch-World trains web agents with 420K verifiable QA tasks — HKUST · 2026-07-21
- GigaAM Multilingual targets low-resource Central Asian ASR with 2M hours of audio — ai-sage · 2026-07-21
- WorldCupArena benchmarks language models on 104 football matches — Zhaokai Wang · 2026-07-21
- Reddit asks whether LLMs need a benchmark for treasure-hunt style reasoning — StrangeOops · 2026-07-21