Two $115 ex-mining APUs run Qwen 35B at 60 tok/s with 64K context for ~$300

Ok-Breadfruit-3523 · reddit · 2026-09-24

A Reddit user built a budget local LLM rig from two ex-mining BC-250 APUs ($115 each, 27GB combined GPU memory), linked via llama.cpp with Vulkan and RPC over 1Gb Ethernet on Bazzite. The setup runs Qwen3.6-35B-A3B at Q4KM, reaching 60 tok/s with 64K context — roughly $300 total including PSU. The author plans to expand to six boards to try running Qwen 3.8 Flash. A replicable reference for budget local deployment.

Original post →

More from Infra

Infra channel →