User reports Qwen model overthinking despite low reasoning effort setting
Jebbyk1 · reddit · 2026-08-18
A user reports that the new Qwen model consumes excessive tokens on thinking regardless of the reasoning effort setting (low/medium/high) in LM Studio. The user suspects a bug where the parameter is not being sent correctly to the LLM.
More from Models
- LiquidAI releases LFM2.5-VL-3B-WebGPU, a multimodal model optimized for browser deployment — LiquidAI · 2026-08-18
- McByte Sets New SOTA on SportsMOT Benchmark — NielsRogge · 2026-08-18
- Benchmark: Qwen 3.8 27B hits 50-60 t/s on dual RTX 5060 TI cards — chocofoxy · 2026-08-18
- Qwen 3.8 27B vs DeepSeek Flash: Benchmaxing or Real? — Best_Sail5 · 2026-08-18
- User claims DeepSeek V4 Pro is mis-trained: cheaper Flash beats it — karminski3 · 2026-08-18
- Qwen3.8-27B Uncensored MLX build trends on Hugging Face for Apple Silicon — orcarouter · 2026-08-18