Qwen 3.8 users report better results forcing reasoning effort to low, llama.cpp flag shared
Jorlen · reddit · 2026-09-03
Reddit users discuss whether Qwen 3.8's (flash next and 27b) default extra-high reasoning effort is counterproductive, with many reporting better results forcing it to low. The OP shares the llama.cpp flag --chat-template-kwargs '{"reasoningeffort":"low"}' and asks how others configure it.
More from coding & agent
- Hamel Husain: everyone is building this AI agent — don't — hugobowne · 2026-09-03
- Dev accidentally burns $100 in an instant running ultracode AI coding mode — zsakib_ · 2026-09-03
- Hidden-bug eval across 105 issues: Fable 5.1 finds 43, none fixes all — cost per model compared — PawelHuryn · 2026-09-03
- Indie author builds an agent-native distribution layer for his novel with A2A endpoints — patternflow · 2026-09-03
- Bezalel gives AI agents memory, email, money and a cloud desktop behind one MCP URL — Rasmic · 2026-09-03
- jjk-explain turns any concept into a Jujutsu Kaisen-style explainer video with one Claude Code command — teortaxesTex · 2026-09-03