Qwen3.8 QAT Q2 review: Good for quick answers, fails at long context
kirisoraa · reddit · 2026-08-29
User tested the Qwen3.8-27B-QAT-Q2 model. It performed surprisingly well on short tasks (e.g., finding documentation bugs). However, during long-context tasks (50k output), the model failed significantly: it started outputting thinking blocks into tool calls, causing validation errors and ineffective loops. The conclusion is that it's suitable for quick answers but unusable for long-context workflows.
More from Models
- MiniMax H3 Max Criticized for Extreme Censorship on Fal — MrUtterNonsense · 2026-08-29
- User Predicts Qwen-4-27B Will Be a Game Changer — Steus_au · 2026-08-29
- AlayaWorld Tops World Model Rankings as First Open-Weights Leader — Obvious_Set5239 · 2026-08-29
- Terminal-Bench 4.0: GLM-5.3 Surpasses GPT-5.6 in New Ranking — eyishazyer · 2026-08-29
- Minimax H3 limits: Identity leaking and quality degradation — Kooky-Mode3047 · 2026-08-29
- Qwen3.8 27B runs at 50 tok/s with 100k context on 16GB GPU — qaf23 · 2026-08-29