ICML Research: Proof Verification Scaling and Long-Horizon Planning

VectorInst · x · 2026-07-08

Other ICML research highlights this week include: Renjie Liao on scaling mathematical proof verification; Florian Shkurti on stable offline-to-online reinforcement learning transfer; and Igor Gilitschenski on test-time planning, which successfully improved the success rate of long-horizon tasks from near zero to over 90%.

Related event: Research Unravels GRPO Training Collapse; ICML Highlights AI Scaling(2 posts)→

Original post →

More from Research

Research channel →