Qwen 3.8 may drop today as 27B dense model, killing MoE offloading speed

zyxciss · reddit · 2026-08-14

Reddit users speculate Qwen 3.8 could be released today as a 27B dense model. If so, it would abandon MoE sparsity, potentially making local inference much slower. Previously, Qwen 3.6 35B-A3B achieved 70 tok/s on an RTX 3060, but dense offloading could be 100x slower.

Original post →

More from Models

Models channel →