Qwen 3.8 vs 3.6: Low reasoning mode loops less
Lair98 · reddit · 2026-08-22
User reports that Qwen 3.8 loops significantly less than 3.6 when reasoning is set to 'low', even on 3-bit quantization.
Key findings:
- The model defaults to xhigh; users must specify 'low' to reduce looping.
- A new parameter preservethinking allows keeping/discarding the reasoning process after each turn. Activating this prevents the model from re-reasoning the same content.
- Real-world testing shows 3.8 'low' mode is an improvement over 3.6, though not perfect.
More from Models
- Claude Security Now Powered by Mythos 5 for Cross-File Vulnerability Scanning — claudeai · 2026-08-22
- Hands-on with Ox Alpha: Impressive Performance in Pi Harness — omarsar0 · 2026-08-22
- Model Self-Talk Artifacts Linked to Synthetic Data Training — ctjlewis · 2026-08-22
- Ornith 1.5 35B live on RunInfra: 262K context, ~$0.02/1M effective input with cache — alejandroll10 · 2026-08-22
- Developer doubts Ox-alpha performance, suspects marketing stunt — bindureddy · 2026-08-22
- SemiAnalysis Deep Dive: Are Open Models Catching Up to Closed Frontier? — JosephJacks_ · 2026-08-22