MTP Doesn't Speed Up MoE Models? Qwen3.8 27B Gets 2x Boost, 35B Unchanged

chibop1 · reddit · 2026-08-16

A Reddit user tests Lightning MTP on OMLX and finds it speeds up Qwen3.8 27B (dense) by 2x, but has no effect on Qwen3.6 35B (MoE). Detailed benchmarks are provided, questioning whether MTP is incompatible with MoE architecture.

Original post →

More from coding & agent

coding & agent channel →