Freeform 偏好学习:自然语言轴定义奖励,机器人操作任务提升 38 个百分点

StanfordAILab · x · 2026-09-10

Chelsea Finn 团队等在 RoboPapers 播客介绍了 arXiv 论文《Freeform Preference Learning for Robotic Manipulation》(FPL)。

所属事件:自然语言定义奖励轴,机器人偏好学习大幅提升(2 条相关)→

原文链接 →

「具身」频道最新

更多「具身」频道 AI 资讯 →