Industry Focus: The Massive Compute Behind Frontier Model Post-Training

teortaxesTex · x · 2026-07-03

Researchers are questioning the actual post-training compute of current frontier models (e.g., V4-Pro, around the 1e25 magnitude). Is it possible that another 1e25 of compute was invested in rollups within two months? The compute ratio between pre-training and post-training is becoming an industry focal point. This issue is crucial for understanding the path to improving model capabilities.

Original post →

More from Infra

Infra channel →