Can Qwen Flash Next Run on a 64GB RAM iGPU Mini PC? One User's Experiment

SomeITGuyLA · reddit · 2026-10-03

A Reddit user explores running Qwen Flash Next on a non-Mac mini PC with 64GB unified RAM and a 780M iGPU, noting others have run it with 12GB VRAM + 64GB RAM and on Macs. He currently runs 125B Ling 3.0 Flash at Q2 quants via llama.cpp (vulkan), and proposes offloading ngrams to SSD — unsupported in llama.cpp, and other engines lack vulkan support.

Original post →

More from Infra

Infra channel →