Benchmarking DeepSeek V4 vs Qwen 3.8 on DGX Sparks

Legitimate_Hat_7852 · reddit · 2026-08-27

The author tested DeepSeek V4 0731, Qwen 3.8 Flash, and GLM 5.3 Flash on a cluster of 4 DGX Sparks. Results showed GLM 5.3 was overly verbose and slow (22 tok/s on dual cards), performing poorly. The final setup involves running DeepSeek V4 on 2 cards for planning/building and Qwen 3.8 on the other 2 for exploration/sub-agent work, creating an effective hybrid workflow.

Original post →

More from coding & agent

coding & agent channel →