Qwen3.8 vs 3.6 Benchmark: 3.8 is 5-20% slower but holds quality
Radiant_Condition861 · reddit · 2026-08-15
A benchmark on RTX PRO 6000 Blackwell comparing Qwen3.8-27B-FP8 and Qwen3.6-27B-FP8 using MTP Sweep shows Qwen3.8 is 5-20% slower across all steps. However, quality remains comparable or slightly better in some cases. If raw speed is the priority, 3.6 remains superior, especially at higher MTP steps.
More from Infra
- NVIDIA NeMo Switchyard Enables Routing AI Agents Across Models — NVIDIAAI · 2026-08-15
- NVIDIA open-sources NeMo Switchyard for dynamic model routing in agent workflows — NVIDIAAI · 2026-08-15
- Vercel ranked as the world's fastest AI Gateway infrastructure — cramforce · 2026-08-15
- mcpp: Auto-generate MCP servers from C++ code via reflection — karurochari · 2026-08-15
- RTX 3090 gets 35 t/s on Qwen 3.8 27B — cviperr33 · 2026-08-15
- CME to launch futures contracts tracking Nvidia H100/B100 compute costs — AccBalanced · 2026-08-15