A chess testbed studies how to split compute between pretraining, SFT, and RL

Pavel_Izmailov · x · 2026-07-21

The paper introduces a chess-based testbed to study how compute should be split across pretraining, SFT, and RL.

Related event: New Research Proposes Joint Scaling Law for Pretraining and RL(18 posts)→

Original post →

More from Research

Research channel →