ServeLearnBench: agents still show large learning gaps in continual self-improvement

BeidiChen · x · 2026-10-08

InfiniAI Lab and Beidi Chen introduce ServeLearnBench, testing whether agents can continually self-improve from serving experience, where needed knowledge is hidden and shifts over time. Across 5 learning harnesses and 6 models, key findings: (1) large learning gaps—agents solve tasks when the hidden policy is given but struggle to discover it from experience; (2) no free lunch for adaptation—continual learning can be costly and even degrade already-correct behavior; (3) exploration is a key bottleneck.

Original post →

More from coding & agent

coding & agent channel →