Ascend SuperPOD optimization lifts DeepSeek-V4 post-training MFU to 34.22%

pmttyji · reddit · 2026-07-23

Full-parameter post-training on Ascend SuperPOD reaches 34.22% MFU

A paper on SLAI T-Rex describes an end-to-end optimization stack for full-parameter post-training of the trillion-parameter DeepSeek-V4 family on Ascend NPU SuperPOD infrastructure.

The repo and paper present this as a full-stack path from infrastructure tuning to domain-specialized reasoning models.

Related event: Ascend SuperPOD Optimization Boosts DeepSeek-V4 Training MFU to 34.22%(2 posts)→

Original post →

More from Infra

Infra channel →