Qwen3.8-27B Runs Impressively on AMD Strix Halo
seti_at_home · reddit · 2026-08-17
A local test of Qwen3.8-27B Q80 on a ROG Flow Z13 with Ryzen AI Max+ 395 (128GB unified memory) showed impressive results. The model successfully generated a simple flight simulator using agent tools. Running via llama.cpp ROCm with native MTP speculative decoding, it achieved 9-19 tok/s with 97-99% MTP acceptance rates.
More from coding & agent
- Using Codex for remote cluster control from phone — iScienceLuvr · 2026-08-17
- Agent Decision Logging Framework for Better Observability — blaizedsouza · 2026-08-17
- Dokie AI MCP Integration Automates Brand-Consistent Slide Generation — Aiden_Tech_Ai · 2026-08-17
- Multica open-sources a harness for your harnesses: assign issues to 20 coding agents — jiayuan_jy · 2026-08-17
- Reverse-engineered recipe for AI 3D games: threejs scaffold plus AI-generated .glb models — filiksyos · 2026-08-17
- One command setup: Claude Cowork configures itself via /setup-cowork — rubenhassid · 2026-08-17