d-Matrix partners with NVIDIA to plug Raptor XPUs into NVLink Fusion rackscale systems

bookwormengr · x · 2026-09-12

Inference chip startup d-Matrix announced a multi-year collaboration with NVIDIA to integrate its next-gen Raptor XPUs, built on 3D DRAM stacked memory, into NVIDIA MGX rack-scale systems via NVLink Fusion, targeting ultra-low-latency premium token services for AI labs, hyperscalers, and neoclouds. The rack design spans NVIDIA's Vera CPUs, NVLink switches, BlueField-4 DPUs, ConnectX-9 SuperNICs, and Spectrum-X Ethernet, with Astera Labs providing connectivity solutions. Customers can run Raptor racks standalone or alongside NVIDIA GPUs, matching compute to each workload phase. The deal signals specialized inference silicon entering NVIDIA's ecosystem rather than competing against it.

Original post →

More from Venture

Venture channel →