Qwen3.8-27B community feedback: memory, speed, and creative writing questions
pmttyji · reddit · 2026-08-15
A Reddit user started a discussion asking for community feedback on Qwen3.8-27B, covering issues like chat template, looping, excessive reasoning, and MTP performance. They also requested comparisons with Qwen3.6-27B on memory and tokens/sec, asked about running Q8 quant with 256K context and KV cache in 32GB VRAM, and inquired about creative writing quality.
More from Models
- Orion-16B passes 100B tokens, largest LLM pretrained with decentralized compute — const_reborn · 2026-08-15
- AI model benchmarks: baseline crushed, top models nearly finish course — const_reborn · 2026-08-15
- Qwen3.8 gets Day-0 support from LightSeek, boosting inference performance by 30%+ — Alibaba_Qwen · 2026-08-15
- DeepSeek's inference cost advantage? User says GLM 5.3 hit daily limits while DS handled it easily — teortaxesTex · 2026-08-15
- Grok 4.6 Boosts Performance: Fable-Level Intelligence, Sonnet Speed, Lower Cost — vikvang1 · 2026-08-15
- Qwen3.8-27B-FP8 on GH200: 10 concurrent streaming requests, first token in 10ms — MaziyarPanahi · 2026-08-15