LeRobot Adds Zero-Shot Reward Models
RemiCadene · x · 2026-07-14
LeRobot v0.6 introduces a unified reward models API along with two zero-shot reward functions:
- Robometer: A general-purpose reward model trained on over 1 million trajectories across 21 robot morphologies. It can directly process video and language instructions to assess task progress and success without requiring task-specific training.
- TOPReward: A lighter alternative that directly uses the output probability of an off-the-shelf VLM to determine "True," eliminating the need for dedicated reward weights.
Both components come with annotation scripts that can write frame-by-frame progress curves back into the dataset, facilitating subsequent reward-aware behavioral training.
Related event: LeRobot v0.6 Adds Zero-Shot Reward Model API(2 posts)→
More from coding & agent
- Six underrated AI skills: copywriting, planning, and domain expertise — tomcrawshaw01 · 2026-07-21
- A redesign gets broken into 40 subagents with specs, not prompts — Wattenberger · 2026-07-21
- RubyConf demo compares LangChain and RubyLLM in a two-minute Rails chat build — kieranklaassen · 2026-07-21
- OpenAI ships faster Codex navigation, steadier sidebar and smoother side chats — OpenAIDevs · 2026-07-21
- Google's Gemini 3.5 Flash Cyber targets CodeMender, while Flash-Lite hits 350 tokens/sec — xiaohu · 2026-07-21
- GenMail wants to turn email into an AI agent workspace across Gmail and Outlook — VraserX · 2026-07-21