Rumor: DeepSeek's rumored single-GPU model may have been trained on Ascend
teortaxesTex · x · 2026-09-30
X user teortaxesTex argues DeepSeek would not open-source kernels hitting 99.8% of hardware limits on GEMM and 98% on MegaMoE without training anything usable end to end. He speculates that at least the rumored small single-GPU model may have been born on Huawei Ascend hardware. Unconfirmed speculation.
Related event: DeepSeek's New Model May Be Trained on Huawei Ascend Chips(3 posts)→
More from Infra
- Quantized softmax attention pretraining: only +0.004 nats loss gap at K=16 with the right calibration — illinois · 2026-09-30
- xLLM training infra open-sourced with xattn attention backend and xBridges toolkit — HongyiWang10 · 2026-09-30
- Auto-research loop on 120 B300s finds 40% Kimi K3 inference gain for $9,176 — bookwormengr · 2026-09-30
- Cerebras to bring 'world's fastest inference' to General Compute — beffjezos · 2026-09-30
- 8% of Asia-to-US air freight is now data center parts — 30 full freighters a day — yacineMTB · 2026-09-30
- mradermacher quants get Gemma 26B to 75 tok/s on 2x RTX 4060 8GB — Spiritual_Impress_30 · 2026-09-30