Qwen 3.8-27B Long Thinking Test: Fails to Finish Planning After 50k Tokens
LippyBumblebutt · reddit · 2026-08-17
A user reported performance issues with Qwen 3.8-27B's long thinking mode on a single RX9070 (Q3KS): it failed to finish the planning phase after 50k tokens for a simple CLI task, whereas the previous 3.6 version completed it easily. Adjusting reasoningeffort and KV quantization settings did not resolve the issue.
Related event: Qwen xhigh Mode Consumes Massive Token Count(2 posts)→
More from Models
- Qwen3.8-27B hits 206 tok/s on single RTX 5090 via SGLang — StefanoGogioso · 2026-08-17
- antirez Optimizes DwarfStar: 170 t/s Generation and 22k tokens/s Prefill on Station — antirez · 2026-08-17
- OpenAI Introduces Tiered Access and Launches GPT-5.6-Cyber Security Model — dl_weekly · 2026-08-17
- Alibaba Cloud still offers cheap DeepSeek models — tobowers · 2026-08-17
- Claude Personification Moment: Rejecting Users and Judging Intentions — ctjlewis · 2026-08-17
- Qwen3.8 Benchmarks: MTP Settings Impact Throughput, Q4 Outperforms Q8 — New-Inspection7034 · 2026-08-17