Reddit Guide: Taming Qwen3.8 Overthinking with Task-Scaled Reasoning Budgets

Express_Quail_1493 · reddit · 2026-09-08

A Reddit user shares practical budgets to stop Qwen3.8 overthinking: 512–1,024 tokens for simple scripts, 2,048–4,096 for standard coding and refactoring, 8,192 for AIME-level math and nested debugging, 16,384 for competition-level tasks. Fine-tunes promising less thinking often degrade the model. Setting a --reasoning-budget-message of "ok now" teaches the pattern universally. At 2,048, Qwen3.8-27b solved an issue Gemini failed.

Original post →

More from coding & agent

coding & agent channel →