DeepSeek-V4 post-training on Ascend SuperPOD reaches 34.22% MFU

_akhaliq · x · 2026-07-23

SLAI T-Rex: DeepSeek-V4 post-training on Ascend SuperPOD

The paper describes an end-to-end post-training practice for trillion-parameter MoE models on Huawei Ascend SuperPOD.

Related event: Ascend SuperPOD Achieves 34.22% MFU for DeepSeek-V4 Training(3 posts)→

Original post →

More from Infra

Infra channel →