ZLUDA + ROCm runs CUDA-targeted Windows apps on AMD GPUs at ~3% slower than native
_underlines_ · reddit · 2026-09-14
A Reddit post details running CUDA-targeted Windows applications on AMD GPUs via ZLUDA with ROCm/HIP—notable since AMD officially had to stop open-sourcing work in this direction.
Benchmark result: only about 3% slower than native CUDA, making it a solid drop-in for CUDA-restricted workloads. The author hopes it finds its way into many projects and optimizations locked to CUDA.
More from Infra
- No.2 US law firm Latham & Watkins buys Nvidia servers to fine-tune open models in-house — ai · 2026-09-14
- Andrew Chen's homelab for local AI: 5090 eGPU, dual DGX Spark and a routing plugin — andrewchen · 2026-09-14
- Palantir names Nebius its preferred sovereign AI infrastructure partner — pdamodaran · 2026-09-14
- INT21's SwarmOS runs 10,000 agents to evolve and optimize inference engines — bingxu_ · 2026-09-14
- Flawed Routers Flood University of Wisconsin Internet Time Server (2003) — walrus01 · 2026-09-14
- Harvard CS249r ML Systems Vol II performance chapter covers graph compilation, fusion, CUDA profiling — blaizedsouza · 2026-09-14