Unsloth releases Qwen3.8-27B quantized: NVFP4 1.5x faster, retains 92-97% accuracy

danielhanchen · x · 2026-08-14

Unsloth announced NVFP4 and dynamic GGUF quantizations for Qwen3.8-27B. NVFP4 is 1.5x faster than BF16 while retaining 92-97% top-1% accuracy. UD-IQ2XXS retains 82.5% accuracy at 9GB, 83.5% smaller than BF16 (54.7GB). These quantized models run locally on 17GB RAM, making Qwen3.8-27B the strongest model for its size that can run locally.

Related event: Alibaba Open-Sources Qwen3.8-27B Multimodal Model, Tops Hugging Face Trending(26 posts)→

Original post →

More from Infra

Infra channel →