Small Dense Models Hard to Run? User Says MoE More Practical
Forward_Jackfruit813 · reddit · 2026-08-16
Reddit user ForwardJackfruit813 questions the value of small dense models, arguing that despite 27B parameters, they are slow and difficult to run on weak GPUs and unified memory setups, and quantize poorly. In contrast, MoE models like Qwen 3.6 MoE 35B run on more systems and quantize better. He feels dense models are no longer worth it and wonders if he's missing the point.
More from Models
- Qwen 3.8 27B Beats Codex in Coding Benchmarks: Wins 8/13, Costs 1/3 — tokenbender · 2026-08-16
- DeepSeek Accused of Grey Testing for High Scores, Opus 5 Output Quality Questioned — teortaxesTex · 2026-08-16
- Study: Cross-Version Transfer of Qwen Interpretability Lenses — imstilllearningthis · 2026-08-16
- Rumor: dots3 heavily distilled from DeepSeek V3, scores and multimodality excite — teortaxesTex · 2026-08-16
- Z.ai Delays GLM-5.3 Open Weights After Model Unexpectedly Develops Hacking Capabilities — Justgototheeffinmoon · 2026-08-16
- OpenAI Previews GPT-5.6 Sol Ultrafast Mode: 14x Speed Boost Powered by Cerebras — Justgototheeffinmoon · 2026-08-16