CoreWeave RL Rollouts Hot-Loads Policy Weights Into Live Deployments, ~15x Faster Than Redeploys

_ScottCondron · x · 2026-10-07

CoreWeave launched RL Rollouts (preview) to close the inference-training loop in RL post-training: instead of redeploying per checkpoint, it transfers weight deltas and hot-loads them into a live deployment without restarting inference or touching in-flight requests — about 15x faster than a redeploy cycle.

Related event: CoreWeave Launches RL Rollouts with 15x Faster Weight Hot-Loading(2 posts)→

Original post →

More from Infra

Infra channel →