Tinker Cuts RL Training Prices Up to 70%, Adds GLM-5.3-Flash and DeepSeek-v4.1-Flash
soumithchintala · x · 2026-10-10
Soumith Chintala announced major efficiency improvements to Tinker, his fine-tuning/RL training API, driven by customers scaling up long-context RL — with price cuts up to 70% passed on to users. GLM-5.3-Flash and DeepSeek-v4.1-Flash are also live on the platform for cost-efficient long-context workloads. The cuts signal that long-context RL has become a mainstream demand and training costs are dropping fast.
More from Infra
- Cloudflare acquires Deno, will maintain runtime for only one more year — Simon Willison · 2026-10-10
- How apps scale: 2006 bigger servers, 2016 clusters, 2026 rewrite in Rust — tristanbob · 2026-10-10
- After HA Yellow failure and LLM-assisted eMMC debugging, altryne moves to Omarchy VM — altryne · 2026-10-10
- VidAIo claims AI video compression halves file size vs AWS, could cut Netflix's $1B streaming bill in half — markjeffrey · 2026-10-10
- Joseph Jacks: analog neural nets are going to be huge — your brain already runs them — JosephJacks_ · 2026-10-10
- Baseten launches Project Beacon, partners Goodfire for in-line open-model safety monitoring — baseten · 2026-10-10