Depth-Pruned Qwen3.8-27B Released: 22.7B Parameters
peplo1214 · reddit · 2026-08-20
The author depth-pruned the Qwen3.8-27B model down to approximately 22.7B parameters without fine-tuning. The model performs well in coding, agentic use, and multi-turn chat, offering a smaller footprint and faster inference. BF16, q8, and q4 versions are available.
More from Models
- Models can now code entire complex software in under an hour — BLUECOW009 · 2026-08-20
- Upcoming benchmark: Local deployment comparison of Qwen3.8, Gemma4, and GPT-OSS — karminski3 · 2026-08-20
- Claude Code Faces Rough Month; Alternative Models Evaluated — Hesamation · 2026-08-20
- Using Ling 3.0 Tiny as an auxiliary model for Qwen agents — My_Unbiased_Opinion · 2026-08-20
- Luna Remains Most Cost-Efficient, Qwen 3.8 27B Ranks Fourth — MikePFrank · 2026-08-20
- GLM-5.3 hits frontier range; Cerebras claims 30x GPU inference speed — 创业邦 · 2026-08-20