Huawei compute limits leave 800B-model training far out of reach, industry source says
ShakeelHashim · x · 2026-07-23
A commenter quotes an industry call saying the gap with the U.S. is fundamentally about compute resources, not just talent or model capability.
- Training a model “as large as theirs” would require about 50,000 GB300s or Huawei 950s, or 200,000 cards.
- The speaker says current resources are only enough to do more experiments at around the 10B-active scale, and that training an 800B model is still far off.
- The bottleneck is framed as compute on both sides: cards can’t be bought domestically, and capital investment is lower than the U.S.
- A highlighted line says the problem is “basically unsolvable” right now because Huawei’s output is also limited, and training an 800B model would need 200,000 of Huawei’s newest cards.
Related event: Leaked Call Claims DeepSeek Operates with ~20K H100 GPUs(3 posts)→
More from Infra
- RTK Terminal Compression Cuts Tokens but Leaves Your AI Coding Bill Unchanged — Bartaseth · 2026-09-11
- SF Compute founder: buying compute is 'an absolutely awful experience' right now — IgorCarron · 2026-09-11
- SmolVM open-sources persistent computer infrastructure for agents that outlive chat sessions — aniketmaurya · 2026-09-11
- PyTorch Day Korea 2026 launches first offline conf, CFP closes Sept 13 — PyTorch · 2026-09-11
- Local LLM server dilemma: 4x CMP-170HX (price up 53% in 20 days) vs Mac Studio M5 Ultra — rumboll · 2026-09-11
- llama.cpp lands Flash Attention tuning for RDNA4, big prefill gains on AMD — pmttyji · 2026-09-11