Why Qwen 3.8 27B Isn't Overthinking: Compared with GLM and DeepSeek

sukazu · reddit · 2026-08-17

A Reddit user argues that Qwen 3.8 27B is not actually an "overthinker." While it uses significantly more reasoning tokens than version 3.6, comparisons with other Chinese models like GLM 5.3 and DeepSeek V4 Flash/Pro on the same tasks show similar behavior, where the extra reasoning is necessary. The user suggests that the frustration stems from hardware limitations preventing most users from utilizing a 1M context window at 150 tps. Additionally, setting a reasoning budget can mitigate usage while maintaining quality better than the previous version.

Original post →

More from Models

Models channel →