Tinker Cuts RL Training Prices Up to 70%, Adds GLM-5.3-Flash and DeepSeek-v4.1-Flash

soumithchintala · x · 2026-10-10

Soumith Chintala announced major efficiency improvements to Tinker, his fine-tuning/RL training API, driven by customers scaling up long-context RL — with price cuts up to 70% passed on to users. GLM-5.3-Flash and DeepSeek-v4.1-Flash are also live on the platform for cost-efficient long-context workloads. The cuts signal that long-context RL has become a mainstream demand and training costs are dropping fast.

Original post →

More from Infra

Infra channel →