Zach Mueller to AIPerf-benchmark GLM 5.3 Flash, DeepSeek v4 Flash, Qwen Flash Next on x8 PCIe5 GPUs

TheZachMueller · x · 2026-10-09

Zach Mueller is collecting requests for an AIPerf comparison this weekend, running models on x8 PCIe5 across 4 Pro vs 4 Max-Q hardware, preferring stable NVFP4 quants as the low end. Current candidates: GLM 5.3 Flash, DeepSeek v4 Flash, and Qwen Flash Next. He says the goal is to publish useful baselines ahead of the backplane release.

Related event: Hugging Face Engineer to Benchmark Flash Models on NVFP4 with AIPerf(2 posts)→

Original post →

More from Infra

Infra channel →