Together Compute Launches Provisioned Throughput for Open Models

Together Compute introduced Provisioned Throughput, a new serverless offering providing reserved inference capacity for frontier open-source models. The service guarantees TPM throughput and a 99% availability SLA, billed per token for mission-critical applications.

2026-07-08 ~ 2026-07-09 · 2 related posts