Local Deployment on DGX Spark: Exploring Upgrades Beyond Qwen 3.5 122B
Voxandr · reddit · 2026-08-01
After running the Qwen 3.5 122B model on a DGX Spark, a developer started a discussion to explore next-step upgrade options for local models.
Given hardware constraints, the community currently has several mainstream quantization choices:
- Laguna 2.1 at NVFP4
- DeepSeek v4 at Q2
- Inkling-Small at IQ3
The author is asking other users what models they are currently running and how they compare in performance to the 122B baseline.
More from Infra
- a16z: AI Infrastructure Demand Shows No Signs of Slowing Amid Supply Chain Snags — a16z · 2026-08-01
- Together AI Deep Dive: Autoscaling Endpoints for LLM Inference — togethercompute · 2026-08-01
- Vercel AI Gateway Adds Team and Project Spend Budgets — cramforce · 2026-08-01
- Tesla Signs 469MW Solar Deals to Lock in AI Compute Power Years Ahead — XFreeze · 2026-08-01
- Analyst Spots Equinix Expanding San Jose Campus by ~200MW — BenBajarin · 2026-08-01
- Micro Center Reports Major Price Hikes for Nvidia RTX 5090 — soumitrashukla9 · 2026-08-01