The Mystery of China's Model Training Compute

zephyr_z9 · x · 2026-07-17

The quoted content questions how Chinese labs manage to train a "3T-level" strong model when US frontier labs possess significantly more and stronger compute power.

Possible explanations mentioned include: Huawei training chips might have largely caught up; Chinese labs might have acquired Blackwell chips through other channels; or the actual scale of compute used is much larger than the outside world imagines. Overall, it's speculation centered around training compute, chip generational gaps, and export controls.

Related event: Kimi K3 Triggers a Reassessment of Chinese Frontier AI(94 posts)→

Original post →

More from Infra

Infra channel →