DeepMind EXIMO: VLM Guided Robot Policy Exploration

deepmind · hf · 2026-08-21

DeepMind introduces EXIMO, which efficiently fine-tunes large vision-language-action robot policies by combining VLM-guided exploration, imitation on orchestrated data, and residual off-policy reinforcement learning.

Original post →

More from Embodied

Embodied channel →