Community Tests Defend Qwen 3.8 27B's Heavy Reasoning Token Use

Qwen 3.8 27B was accused of "overthinking" for generating a huge number of reasoning tokens (measured at over 16K in some tests), but on 08-17 several community members pushed back with hands-on tests, arguing that this deep thinking is the necessary cost of performance approaching much larger models—while other evaluations point out its weaknesses in logic and visual reasoning.

Confirmed

Why it matters

2026-08-17 ~ 2026-08-17 · 5 related posts

Primary sources