把模型放进 harness 里训练:LFM2.5 的 agentic RL 细节

helloiamleonie · x · 2026-09-25

Leonie 解释 agentic RL 环节:既然模型如今主要通过 agent harness 使用,就直接在 harness 内训练,让模型提前熟悉各 harness 的 system prompt 和内置工具。整体架构分三块:tasks(写演示文稿、分析报告等办公类 agentic 任务)、harness(模型通过 harness 处理任务)、sandbox(每次 rollout 起一个全新沙箱,模型经 harness 与工具和本地文件系统交互)。

所属事件:Liquid AI 发布 LFM2.5-2.6B 并拆解端侧训练配方(10 条相关)→

原文链接 →

「编程与Agent」频道最新

更多「编程与Agent」频道 AI 资讯 →