Together launches preemptible compute for GPU clusters at 50% of on-demand price
togethercompute · x · 2026-09-11
Together AI announced public preview of preemptible compute for Together GPU Clusters:
- Pricing: same NVIDIA GPU infrastructure as on-demand, flat at 50% of on-demand rate with sub-hourly billing — no spot-market volatility.
- Mechanics: preemptible nodes draw from unused capacity and can be reclaimed. Reclamation triggers up to a five-minute drain window: the node is cordoned, pods receive SIGTERM, and workloads get up to five minutes to checkpoint and exit. The cluster automatically refills toward its preemptible target.
- Use cases: evals, fine-tuning, batch inference, and short interruption-tolerant experiments; standard nodes remain synchronously allocated and never preempted.
- Available today on Kubernetes clusters in all regions, on new or existing clusters.
A practical option for teams trading interruption tolerance for a 50% cost cut.
More from Infra
- k3 Report Section Confirms Millions of Concurrent Sandboxes in Its RL Training Run — stochasticchasm · 2026-09-11
- Pentagon in talks to lend roughly $5 billion to AI cloud startup Fluidstack — vitaliychiley · 2026-09-11
- Eric Schmidt: AI may hit a money wall before a power wall — $1T capital needed — rohanpaul_ai · 2026-09-11
- SpaceX signs another AI compute deal: $1.11B per month, on track for $100B ARR — NinaDSchick · 2026-09-11
- Carmack: Jetson Thor's 128GB at 273GB/s is over-provisioned for real-time robotics — ID_AA_Carmack · 2026-09-11
- YC Demo Day startup touts ultra-pure diamond wafers for data centers, $160M in LOIs — ycombinator · 2026-09-11