Text-Only Qwen3.5 2B/4B/9B MLX 4-bit Packages Released, 2B Is Just 1GB

sachasayan · reddit · 2026-10-09

A Redditor released text-only (vision-removed) MLX 4-bit packages of Qwen3.5 for Apple Silicon local inference: 2B, 4B and 9B, each in original and Huihui abliterated variants — six packs total, with the 2B version at only 1GB. The author runs them in their writing app Minstrel for summarization and contextual decisions; downloads are on Hugging Face, crediting Qwen, Huihui and community converters.

Original post →

More from Models

Models channel →