EmbeddedLLM, lead vLLM ROCm maintainers, finally get persistent AMD MI355X cluster
AccBalanced · x · 2026-09-23
SemiAnalysis gave a shoutout to EmbeddedLLM, the lead maintainers and CODEOWNERS of vLLM ROCm. Until recently the team didn't even have persistent access to an MI355X cluster; only after SemiAnalysis worked with AMD leadership did they finally get a stable cluster to make MI355X great on AMD.
The episode highlights how AMD's inference ecosystem has long lacked sustained compute investment, and how vLLM's AMD support rests on a small third-party team.
Related event: EmbeddedLLM, vLLM ROCm maintainer, finally gets MI355X cluster(3 posts)→
More from Infra
- Unsloth Desktop Hotfix Adds Qwen-Image-2.1 Image Editing and Fixes GGUF Loading — danielhanchen · 2026-09-23
- Qwen 3.6 35B-A3B Q6 hits ~50 tok/s on a 128GB Strix Halo — what's the best local model now? — jankeydankey · 2026-09-23
- Together AI adds canary rollouts for zero-downtime model upgrades on dedicated inference — togethercompute · 2026-09-23
- Dedicated Hardware for Running AI Agents at Scale Arrives — cyrilzakka · 2026-09-23
- Ternary Bonsai 2 27B: 5.9GB weights retain ~95% of full-precision reasoning — cephaloform · 2026-09-23
- Qwen 27B runs 24hr unattended on one RTX5090, builds full Postgres-SpringBoot-React spreadsheet app — anglepoiselife · 2026-09-23