Qwen 3.8 27B vs Qwen Flash Next on an M3 Max: near-identical feel, faster prefill on 27B

Zeeplankton · reddit · 2026-09-07

The poster runs Qwen 3.8 27B and Qwen Flash Next locally on an M3 Max 96GB: both feel largely identical, though 27B prefills faster; they ask whether anyone is improving prefill performance in MLX.

Side question: is anyone building a harness that works with reasoning fully off, inspired by JetBrains revealing Junie runs Qwen 3.6 with reasoning disabled entirely.

Original post →

More from Infra

Infra channel →