Qwen 27B Now Runs on AMD NPUs via FastFlowLM, at a Slow 1 tps
TuskNaPrezydenta2020 · reddit · 2026-10-01
FastFlowLM v1.0.7 now runs the 27B Qwen model locally on AMD NPUs—though the poster jokes the decode speed is just 1 token per second. Open source under ROCm/FastFlowLM on GitHub.
More from Infra
- Swarms Rust claims 130-440x faster startup than LangChain, LangGraph and CrewAI — KyeGomezB · 2026-10-01
- India's EtherealMachine Builds Its Own 5-Axis CNC Machines From Scratch — RoboBalaji · 2026-10-01
- Merge Agent Handler ships as a partner recipe in NVIDIA NemoClaw for safe enterprise agents — shensi · 2026-10-01
- Google's Data Agent Kit hits GA, wiring 15+ data services into your coding agent via MCP — rseroter · 2026-10-01
- The agent loop is what matters: why local LLMs keep you in control — ag789 · 2026-10-01
- Framework opens preorders for AMD Ryzen AI Max 400 desktop with 192GB RAM — Educational_Sun_8813 · 2026-10-01