Unsloth V3 Qwen Model Broken on Dual AMD GPUs, V2 Fix Available
Equivalent-Ear-8016 · reddit · 2026-08-21
A user reported that the Unsloth V3 version of the Qwen3.8-27B quantized model crashes on a dual AMD GPU setup (RX 9070 XT + RX 7800 XT) running Windows with Vulkan. The issue occurs immediately upon generating the first token. Testing revealed that rolling back to the previous V2 version (revision 408fcc1807ab) resolves the stability issue. The user provided a direct link to the working Q4KM V2 GGUF file for others affected by the bug.
More from Infra
- 456 tok/s Qwen 3.8 on Modded RTX 2080 Ti via NInfer Port — xrailgun · 2026-08-21
- Helium Mirrors Removed from Codeberg Amid Criticism of Policy — uwukko · 2026-08-21
- Optimizing Qwen3.8-27B: From 9.5 to 153 Tokens/Second — TrifleHopeful5418 · 2026-08-21
- Nvidia shorted over cuBLAS not using asymptotically optimal algorithms — basedjensen · 2026-08-21
- Fastest Qwen 3.8 27B version for AMD GPU? — Gloomy_Letterhead395 · 2026-08-21
- CUHK Team Open Sources Libra: 3x Throughput for Agentic Training — jiqizhixin · 2026-08-21