Dual DGX Spark Setup Hits 40 tok/s on DeepSeek Locally
Teknium · x · 2026-08-09
Developer Teknium shared real-world benchmarks for local AI deployment using NVIDIA DGX Spark. By connecting two Spark units with a single cable, he achieved around 40 tok/s running an uncensored DeepSeek model without extra acceleration frameworks. This enables completely private, local inference. He also expressed hope that the upcoming Spark 2 will feature 512GB of memory.
Related event: Dual Spark Runs Abliterated DeepSeek Locally at 40 tok/s(3 posts)→
More from Infra
- LifeOS: A Local, Voice-Driven Personal Organizer — Extension-Bid-639 · 2026-08-24
- Hyperscalers: Choosing Between HDD and SSD Based on Space and Cost — generativist · 2026-08-24
- Samsung shows new HBM cooling solution, hints at die performance variance — BenBajarin · 2026-08-24
- Tobi open-sources walgit: A single-binary Git server backed by object stores — jevon · 2026-08-24
- s3collections: Durable Go data structures backed directly by S3-compatible storage — andersonbcdefg · 2026-08-24
- Prediction market gives 68% chance of a state data center moratorium by year-end — Polymarket · 2026-08-24