Report: DeepSeek training a 2T-param model with 8T planned, Huawei chips due late 2026
Hesamation · x · 2026-09-21
According to a report, DeepSeek is training a 2T-parameter model and has an 8T model planned. For context, Kimi K3 is 2.8T and DeepSeek V4 Pro is 1.6T.
The company also reportedly expects new Huawei training chips in Q4 2026 or Q1 2027, with the CEO telling investors that moving to domestic-chip training “has to work.” Note: unverified rumor.
Related event: DeepSeek Reportedly Training 2T-Parameter Model, 8T Planned(2 posts)→
More from Infra
- Colibrì's Brio mode scores candidate answers instead of generating text locally — Just_Vugg_PolyMCP · 2026-09-21
- Rothschild Redburn goes bearish on AI compute, rates NBIS and CRWV Sell — JOBhakdi · 2026-09-21
- PC memory module prices keep climbing daily as Meta's Muse holds #1 on the App Store — firstadopter · 2026-09-21
- Quantizing Cellpose-SAM for stem cell imaging: W4/W8 hits 6.76x compression with zero failures — capicu-ai · 2026-09-21
- CXMT's wafer capacity to grow 10x by 2030, nearing Micron parity in late 2026 — Terminator857 · 2026-09-21
- Building a 4-GPU local LLM rig: why Threadripper beats LGA1700 on PCIe lanes — El_90 · 2026-09-21