320 H100s for three months: a checkpoint that serves as a compute warning

teortaxesTex · x · 2026-10-06

teortaxesTex highlights a research post stating all its results come from a checkpoint trained from scratch on just 320 H100 GPUs over three months—a tiny fraction of frontier language and video model compute.

He frames this as a warning/research note, estimating 691K H100-hours at roughly 2e24 fp8 FLOPs.

Related event: Checkpoint Trained on Just 320 H100s Sparks Compute Threshold Debate(2 posts)→

Original post →

More from Infra

Infra channel →