More compute shifts the optimum from pretraining toward RL in a chess scaling law

Pavel_Izmailov · x · 2026-07-21

The paper argues that the optimal allocation of compute between pretraining and RL depends on total budget.

Related event: New Research Proposes Joint Scaling Law for Pretraining and RL(18 posts)→

Original post →

More from Research

Research channel →