SGLang talk says RL post-training is really an inference problem

BanghuaZ · x · 2026-07-24

A talk at a dstack / Crusoe / SGLang event argues that reinforcement learning post-training is fundamentally an inference problem.

Original post →

More from coding & agent

coding & agent channel →