DeepSeek V5 leak: 2T parameters, reportedly trained fully on Huawei Ascend chips
teortaxesTex · x · 2026-09-27
An unverified leak claims DeepSeek is preparing an imminent V5 launch: rumored at 2 trillion parameters (not 3T), described by founder Liang Wenfeng as the company's biggest bet yet. It would reportedly be the first DeepSeek model trained fully on Huawei Ascend chips instead of Nvidia — though needing roughly 4x more Ascend accelerators to hit the same training scale. The open-weight strategy is said to continue. A joked-about "independent corroboration" has surfaced from Nigeria. All rumors, unconfirmed.
More from Infra
- Harvard Puts Full ML Systems Curriculum CS249r Online for Free — techNmak · 2026-09-27
- SpaceX's vertical empire: from rockets to Colossus AI compute powering Grok — XFreeze · 2026-09-27
- Token-efficient reasoning model Swift-1.5-Qwen3.8-27B trends on Hugging Face — ukisai · 2026-09-27
- Hacker uses NVIDIA DGX Station to run local models for CAD, invites SF folks to join — hudzah · 2026-09-27
- Yacine goes 'absurdly bullish' on TensorTorrent in rare endorsement — yacineMTB · 2026-09-27
- Pathway's 150M-parameter BDH reasons in latent space, beyond Transformers — bigdata · 2026-09-27