Inference startup: training wins inference workloads, models should improve with use

ypatil125 · x · 2026-09-23

An inference provider shares months of customer-facing lessons: the business took off, but training wins and keeps inference workloads. Inference isn't a commodity when you can improve the model itself — using online training like On-Policy Self Distillation and frontier-grade RL post-training to target specific behaviors against customer evals. The thesis: models should get better the more you use them.

Original post →

More from Infra

Infra channel →