Alibaba's Qwen3.8-Max runs autonomous chip design for 60 hours, cutting power 59.5%

千问大模型 · wechat · 2026-09-22

Qwen announced two production-grade "chip-model co-evolution" experiments with Qwen3.8-Max. In chip design, given only a NoC module spec, the model autonomously ran 60+ hours with 10k+ tool calls through RTL, self-built verification, and physical implementation — cutting standard cells 29%, area 42%, and power 59.5%. In inference adaptation, it ported Qwen3.8-Flash-Next to the T-Head Zhenwu M890 in 56 hours with 40+ sub-agents, halving TTFT (−47%), cutting TPOT 60%, and boosting single-instance throughput 96%, driven by 8 self-written kernels and CUDA Graph capture, with bit-exact outputs. Humans still set goals, acceptance criteria, and sign-off.

Original post →

More from coding & agent

coding & agent channel →