Together Launches Provisioned Inference Capacity

togethercompute · x · 2026-07-08

Together Compute has introduced Provisioned Throughput, offering reserved inference capacity for frontier open-source models. It features token-based billing and guarantees a 99% availability SLA. The company claims the solution combines serverless ease-of-use with more stable capacity guarantees, reducing costs by up to 90% compared to Opus 4.8.

Related event: Together Compute Launches Provisioned Throughput for Open Models(2 posts)→

Original post →

More from Infra

Infra channel →