Next-gen model training could hit 5x pretraining compute if 300k GB200 rumor holds

scaling01 · x · 2026-09-07

scaling01 estimates next-gen frontier model training compute: 1-2x pretraining with 100k GB200s, but closer to 5x if the rumored 300k GB200s materialize. On top of that, RL may add another 2-5x of pretraining-scale compute.

Related event: KOL Tests Astra, Sees No AGI; Speculates GPT-6 Uses Up to 5x Compute(3 posts)→

Original post →

More from Infra

Infra channel →