Researchers split on whether scaled self-play with capped APM can break model ceilings

Kangwook_Lee · x · 2026-09-26

Rémi Leblond argues a game AI trained purely via self-play from scratch, no human data, would improve further if capped at 500 APM peak / 300 APM average and trained with 10x more compute. Prof. Kangwook Lee is skeptical: unlikely, with no evidence supporting it — citing AlphaStar. A concise disagreement on whether scaled self-play keeps paying off.

Related event: Self-Play StarCraft AI Sparks AGI Benchmark Debate(4 posts)→

Original post →

More from Research

Research channel →