"DeepSeek trained R2 on Ascend" rumor is 90% fake, argues AI researcher citing model-line evidence
NunoSempere · x · 2026-09-25
Responding to reporting that DeepSeek tried to train its R2 model on Huawei Ascend chips, teortaxesTex lays out why he thinks it's a broken-telephone rumor:
- He's 90% sure no "R2" project ever existed — DeepSeek unifies model lines and serves as few checkpoints as possible, so DS-R was bound to be folded in like DS-Coder
- Why test a flagship on a new stack instead of a V3.5-lite? Zhipu trained a 9B GLM-Image on Ascends but still runs flagships on Nvidia
- The best Huawei hardware then was the weak 910C with terrible software support; the only recent large model on such hardware came from Meituan
- NunoSempere notes the "DS never trained on Huawei" conclusion likely holds, but it mainly says something about Epoch's evidence standards
Related event: Blogger disputes report that DeepSeek trained R2 on Huawei Ascend(2 posts)→
More from Infra
- Investor Predicts EDA/CAD Will Collapse Into One Flow Within 3-5 Years — ai · 2026-09-25
- AI Data Center Debt Starting to Roll Over, Rising Rates Accelerating the Problem — AIFlow_ML · 2026-09-25
- Musk details xAI compute: Colossus 2 to hit 880k GB300s by year-end — elonmusk · 2026-09-25
- Qwen-Image-2.1 gets GGUF quantization, could run text-to-image on a Snapdragon 865 phone — ResidentAping · 2026-09-25
- Deep Inference-Query Engine Integration: Custom Scheduler and Workload-Aware KV Cache for Prefill-Only AI Filters — charles_irl · 2026-09-25
- Finance worker seeks local AI setups to cut soaring Codex/ChatGPT costs — Startup__Sam · 2026-09-25