Qwen 3.8 users report better results forcing reasoning effort to low, llama.cpp flag shared

Jorlen · reddit · 2026-09-03

Reddit users discuss whether Qwen 3.8's (flash next and 27b) default extra-high reasoning effort is counterproductive, with many reporting better results forcing it to low. The OP shares the llama.cpp flag --chat-template-kwargs '{"reasoningeffort":"low"}' and asks how others configure it.

Original post →

More from coding & agent

coding & agent channel →