Scaling in Inference and RL Training Outperforms Pre-training

tobyordoxford · x · 2026-07-03

Toby Ord notes that by standard AI scaling law metrics, current scaling performance is quite solid: halving the error rate in pre-training requires about 1,000,000x the compute. However, on math benchmarks, both inference scaling and RL training deliver logarithmic improvements. This is the most informative post within the same AISecurityInst discussion.

Original post →

More from Infra

Infra channel →