Inference is turning GPU compute into a tradable commodity

ArtificialAnlys · x · 2026-09-09

jessiedong argues that inference is already making compute tradable: inference providers are "partly compute traders" — they reserve compute, run models on it, and sell output by the token, profiting by buying compute cheaply and squeezing more tokens per GPU while taking idle-capacity risk.

Providers turn disparate hardware into an easily comparable commodity (the same model's tokens at set price and speed). The market already shows a division of labor:

If you're bearish on trading GPUs, look at inference instead.

Original post →

More from Venture

Venture channel →