A joint scaling law links RL performance to model size, pretraining tokens, and compute

Pavel_Izmailov · x · 2026-07-21

The paper proposes a joint scaling law for RL performance in terms of model size N, PT tokens T, and RL compute C.

Related event: New Research Proposes Joint Scaling Law for Pretraining and RL(18 posts)→

Original post →

More from Research

Research channel →