Breaking the AI Memory Wall: CXL Nears Commercial Deployment

BenBajarin · x · 2026-08-10

Analyst Ben Bajarin highlights that as AI shifts from basic training to advanced inference with long context and RAG, demand for HBM and enterprise SSDs will peak again. However, memory capacity within CPU/GPU packages is hitting physical limits, known as the "memory wall."

He argues that CXL (Compute Express Link) is the key to breaking this bottleneck, allowing memory pooling and expansion via PCIe. While deployment architectures are still maturing, CXL is expected to begin commercial deployment next year in custom hyperscaler clusters, scaling into 2028. Solving the "fleet-level" memory allocation for AI workloads like KV cache is becoming an urgent industry priority.

Original post →

More from Infra

Infra channel →