Distilled Qwen3.8-35B-A3B GGUF quantized model trends on Hugging Face

empero-ai · hf · 2026-09-21

empero-ai's Qwen3.8-35B-A3B-Distill GGUF model is trending on Hugging Face. The distillation-based reasoning model uses a MoE architecture (35B total, 3B active params) with gated-deltanet layers, and runs via llama.cpp.

Original post →

More from Models

Models channel →