Running Qwen3.6 on a 16GB Mac

andreaborio · hn · 2026-07-18

Demonstrates how to run Qwen3.6-35B-A3B on a 16GB M1 Pro. The core idea is using SSD-streamed MoE to stream model weights and experts on-demand from the SSD, rather than loading them entirely into memory at once.

Key takeaways from this approach:

Original post →

More from Infra

Infra channel →