CoreWeave brings up multi-rack NVIDIA Vera Rubin NVL72 clusters, shares engineering details
OnlineInference · x · 2026-09-17
CoreWeave announced the bring-up of multi-rack NVIDIA Vera Rubin NVL72 clusters — each rack packs 72 Rubin GPUs, 36 Vera CPUs, NVLink 6 fabric, and 45°C liquid cooling, with hundreds of GPUs networked into one scale-out system for training, inference, and agentic AI workloads. The blog details multi-rack engineering pitfalls: straggler GPUs, marginal network links hurting collectives, cooling anomalies, and config consistency across NVLink domains.
Related event: CoreWeave Lights Up Multi-Rack NVIDIA Vera Rubin NVL72 Clusters(2 posts)→
More from Infra
- How to run Qwen3.8-Flash-Next with N-gram SSD streaming in llama.cpp? — Ambitious_Fold_2874 · 2026-09-17
- How Bell Labs Missed the Microchip: IEEE Spectrum Revisits a Landmark Tech-History Blunder — ArtificialOther · 2026-09-17
- Agentic AI systems are the next network users: 40% of enterprise apps to include agents by 2026 — seankinneyRCR · 2026-09-17
- B200 spot rental up 80% in 8 months as demand outpaces compute buildout — JOBhakdi · 2026-09-17
- MLX-Serve 26.9.3 ships: Qwen Flash Next tops 100 tok/s on M4/M5 Max Macs — TheMoonMidas · 2026-09-17
- Running a 14B Model on 16GB RAM: 'My PC Is a Toaster Now' — Aggravating_Site381 · 2026-09-17