Qwen3.8-27B Runs Impressively on AMD Strix Halo

seti_at_home · reddit · 2026-08-17

A local test of Qwen3.8-27B Q80 on a ROG Flow Z13 with Ryzen AI Max+ 395 (128GB unified memory) showed impressive results. The model successfully generated a simple flight simulator using agent tools. Running via llama.cpp ROCm with native MTP speculative decoding, it achieved 9-19 tok/s with 97-99% MTP acceptance rates.

Original post →

More from coding & agent

coding & agent channel →