Back-of-Envelope: Frontier US Model Estimated at 8e26 FLOPs, 800B Active Params, 170T Training Tokens

teortaxesTex · x · 2026-09-15

TeortaxesTex extrapolates from Liang Wenfeng's May figure about the largest model known to be in development by Americans: assuming fp4 precision on Blackwells, 8.17e26 FLOPs, 800B active params, and a reasonable 4%-5% sparsity (16-20T total params), such a model would need roughly 170T training tokens. He adds that OpenAI can already access 100T+ tokens (citing glm-oss-related data). Speculative but source-anchored estimate of frontier model scale and training resources.

Original post →

More from Infra

Infra channel →