Qwen3.8-Flash-Next on a single R9700: 863 t/s prefill and 35 t/s decode at 230k context

Designer_Elephant227 · reddit · 2026-10-02

A Redditor shared full details of running Qwen3.8-Flash-Next locally on a single AMD R9700 via the exllamav3 rocm fork, asking whether the setup is optimal.

Original post →

More from Infra

Infra channel →