Users report random stop behavior in Qwen 2.5/3 during long context generation
T_rex2700 · reddit · 2026-08-15
Users have reported an issue where Qwen 2.5/3 models stop generation randomly during long context tasks. Even with a 64K context window, the model tends to halt after processing about 3.5K tokens and generating 4K. Attempts to fix the chat template or jinja template have failed, leading to speculation about configuration issues or penalties.
More from Models
- Anthropic's internal model scores leak: new model may surpass Mythos 5 — scaling01 · 2026-08-15
- Qwen 3.8 27b Rivals Opus 4.6? Local Models Challenge Frontier AI, Reddit Debates — Odd_Tumbleweed574 · 2026-08-15
- FrontierSWE Rankings: GLM 5.3 Takes #2, Grok 4.6 #3, Claude Fable 5 Still Top — dejavucoder · 2026-08-15
- OpenAI's o3 Model Non-Functional for Paying Users; User Organizes Complaint Campaign — Aine_123 · 2026-08-15
- Arcee Adds deepseek-v4-pro-0813 to Open Models API — latkins · 2026-08-15
- Three major AI releases: Qwen3.8-27B, GLM-5.3, DeepSeek V4-Pro — johnseach · 2026-08-15