Qwen3.8-27B identical architecture to 3.6, gains purely from training
Course_Latter · reddit · 2026-08-15
A Reddit user notes that Qwen3.8-27B has exactly the same architecture as Qwen3.6-27B (0 changes), implying all capability gains come from training improvements.
More from Models
- Orion-16B passes 100B tokens, largest LLM pretrained with decentralized compute — const_reborn · 2026-08-15
- AI model benchmarks: baseline crushed, top models nearly finish course — const_reborn · 2026-08-15
- Qwen3.8 gets Day-0 support from LightSeek, boosting inference performance by 30%+ — Alibaba_Qwen · 2026-08-15
- DeepSeek's inference cost advantage? User says GLM 5.3 hit daily limits while DS handled it easily — teortaxesTex · 2026-08-15
- Grok 4.6 Boosts Performance: Fable-Level Intelligence, Sonnet Speed, Lower Cost — vikvang1 · 2026-08-15
- Qwen3.8-27B-FP8 on GH200: 10 concurrent streaming requests, first token in 10ms — MaziyarPanahi · 2026-08-15