Too RAM-Hungry? Devs Discuss Best SLMs to Run Locally on 16GB Machines

elie2222 · reddit · 2026-08-12

With recent model releases, running AI smoothly on standard machines (e.g., 16GB RAM) is a hot topic. The author notes that while the 30B parameter class is growing, it's too resource-intensive for average local deployment.

Currently, lightweight models like Gemma 4 e4b/e2b strike a good balance between resource consumption and performance. The upcoming Aion Instruct from Microsoft is also highly anticipated. The community is debating whether Small Language Models (SLMs) will see the same rapid progress as larger models.

Original post →

More from Infra

Infra channel →