Running a 100B+ Qwen3 model locally on 64GB RAM: good vibes, short 192K context

lxfater · x · 2026-10-04

The author deployed a quantized 100B+ parameter Qwen3 model locally on a 64GB RAM machine. First impressions are positive — the new UI looks noticeably better — with the only drawback being the 192K context window. They're now looking for real use cases for their new local powerhouse.

Original post →

More from Infra

Infra channel →