Running Qwen3.8-Flash-Next GGUF locally on an M4 Pro 48GB Mac, dense 27B still faster

JLeonsarmiento · reddit · 2026-09-28

A user reports running Qwen3.8-Flash-Next on an M4 Pro 48GB Mac via the ISTA-DASLab GGUF quantization on Hugging Face. Their take: the dense 3.8-27B is actually faster — and possibly better due to lighter quantization.

Original post →

More from Infra

Infra channel →