Training and Inference Share One GPU Fleet, but Fungibility Only Runs One Way

robleclerc · x · 2026-09-26

A widely endorsed analysis of compute economics argues that while training and inference pull from the same GPU fleet, fungibility is one-way: frontier post-training needs coherent, well-interconnected clusters while inference doesn't, so training can always pull capacity from serving but serving can never backfill training.

Key points:

Original post →

More from Infra

Infra channel →