Distilled Qwen3.8-35B-A3B GGUF quantized model trends on Hugging Face
empero-ai · hf · 2026-09-21
empero-ai's Qwen3.8-35B-A3B-Distill GGUF model is trending on Hugging Face. The distillation-based reasoning model uses a MoE architecture (35B total, 3B active params) with gated-deltanet layers, and runs via llama.cpp.
More from Models
- $5, 10-minute SFT on Qwen3.6-35B-A3B lifts GPQA +8% and MMLU-Pro +12% — josh_wills · 2026-09-21
- Users fight Claude's return to mannered prose and overly safe outputs — ColleenMBrady · 2026-09-21
- How could a model spontaneously write jailbreak language? Reddit probes OpenAI's injection report — sivadneb · 2026-09-21
- GLM 5.3 Flash, DeepSeek V4.1 Flash and Qwen 3.8 all beat 'frontier' models from just 10 months ago — burny_tech · 2026-09-21
- Reddit meme: Grok 'devolves' while local MiniMax 3 wins users over — Mystvearn_ · 2026-09-21
- Two years after 'intelligence too cheap to meter', $10/$50 models are the new norm — teortaxesTex · 2026-09-21