China’s AI stack is converging on MoE and larger supernodes

teortaxesTex · x · 2026-07-20

The post argues that China’s AI ecosystem has converged on MoE plus wide expert parallelism to make models run on weaker NPUs and with lower HBM bandwidth demands.

It also points to a broader hardware trend: system vendors are now designing much larger “supernode” domains. The attached image highlights several Chinese infrastructure demos, including Biren’s 1024-card scale-up cluster with optical interconnect, MetaX’s next-gen AI SuperNode, SUGON’s 8000 SuperCluster for training/inference/scientific computing, and Alibaba’s Zhenwu-890 / Panjiu AL128 SuperNode with 128 chips per cabinet and high-bandwidth interconnect.

Original post →

More from Infra

Infra channel →