Qwen 3.8 27B Vision MLX Quants: Low-Bit Performance Benchmarks

Top-Eye-8104 · reddit · 2026-08-26

The team released various MLX quantizations of the Qwen 3.8 27B Vision model (ranging from 8bit to 3.23bpw) and benchmarked them against community versions like lm-studio and mlx-community. Tested on an H200 with a custom dataset, their 11.8 GB DWQ quant achieved 70.32% top-1 agreement, significantly outperforming others of similar size. This version runs on a 16GB MacBook with a wired limit tweak, utilizing distillation from a BF16 teacher model.

Original post →

More from Infra

Infra channel →