Frontier pre-training runs likely capped at ~2 months to avoid wasting algorithmic progress

scaling01 · x · 2026-09-07

A case that frontier pre-training runs are now capped at roughly two months: spending 100 days on 100K GB200s makes little sense, and long runs lock you out of rapid algorithmic progress — effectively wasting 100 days of it. The author takes this as the base case for GPT-6 Astra's pre-training.

Original post →

More from Infra

Infra channel →