DeepSeek V4 hits 67 t/s on dual GX10 GPUs

koalfied-coder · reddit · 2026-08-30

User reports achieving over 65 tokens/s sustained inference with DeepSeek V4 on dual GX10 GPUs, highlighting the utility of the 2570 context length evaluation.

Original post →

More from Infra

Infra channel →