Qwen3.8-9B Distill Model Gets GGUF Release for Local Use

A GGUF-quantized distill of Qwen3.8-9B has appeared on Hugging Face, supporting llama.cpp for local deployment, with tags hinting at a gated-deltanet architecture.

2026-08-19 ~ 2026-08-19 · 2 related posts