Benchmarking DeepSeek V4 vs Qwen 3.8 on DGX Sparks
Legitimate_Hat_7852 · reddit · 2026-08-27
The author tested DeepSeek V4 0731, Qwen 3.8 Flash, and GLM 5.3 Flash on a cluster of 4 DGX Sparks. Results showed GLM 5.3 was overly verbose and slow (22 tok/s on dual cards), performing poorly. The final setup involves running DeepSeek V4 on 2 cards for planning/building and Qwen 3.8 on the other 2 for exploration/sub-agent work, creating an effective hybrid workflow.
More from coding & agent
- JetBrains Releases Guidelines to Help AI Agents Write Modern Go Code — JetBrains · 2026-08-27
- Claude-mem Provides Persistent Context Across Sessions for Every Agent — thedotmack · 2026-08-27
- llama.cpp Underuses Your NVMe Array? The Author Lets Kimi Tweak the Code — carrigmat · 2026-08-27
- Agno Launches: A 'Self-Building' Agent Platform for the Cloud — pritisinghhhh · 2026-08-27
- OpenCode Senses: Local Vision Plugin with 13 Tools — RevolutionaryPen4661 · 2026-08-27
- Agent Opens Bambu Handy App on Phone to Reprint Job — haydendevs · 2026-08-27