Agentic RL inside the harness: how LFM2.5 trains with sandboxes and tools

helloiamleonie · x · 2026-09-25

Leonie details the agentic RL stage: since models are mostly used through agent harnesses, train them inside the harness so they're already familiar with system prompts and built-in tools. Components: agentic office tasks (presentations, analysis reports), the harness the model works through, and a fresh sandbox per rollout with tool and filesystem access.

Related event: Liquid AI Releases LFM2.5-2.6B and Details Its On-Device Training Recipe(10 posts)→

Original post →

More from coding & agent

coding & agent channel →