Having Enough Compute Doesn't Guarantee a Top-Tier Model

teortaxesTex · x · 2026-07-14

This discussion suggests that players like DeepSeek don't find it hard to secure the compute needed for a "10T" scale pretraining model, as the total volume required isn't as massive as imagined.

The real challenge isn't "can it be trained," but rather:

Related event: Post-Training, Not Compute, is the Real Bottleneck for Top AI Models(2 posts)→

Original post →

More from Infra

Infra channel →