Q-Planning enables recursive self-improvement for robot policies: 25% to 80% success
animesh_garg · x · 2026-08-28
Q-Planning is a learning-based harness that allows large black-box robot policies to recursively self-improve. On a fine-grained task, it boosts success rates from 25% to 80% in 30 minutes (100 robot attempts) without extra human data.
More from Embodied
- MIT's Loop Closure Grasping Balances Strength, Gentleness, and Versatility — lukas_m_ziegler · 2026-08-28
- Stewart Alsop to host in-person robotics workshop in Mendoza — StewartalsopIII · 2026-08-28
- Open Source Animacy: Maps Any Human Video to Robot Motion Without Retraining — dee_hw · 2026-08-28
- Agent Body Protocol: Visualizing Coding Agent Status — dee_hw · 2026-08-28
- Autonomous Lamp: Open Source AI Companion with Physical Body — dee_hw · 2026-08-28
- Google Introduces AgentHands: Gesture-Enabled XR Conversational Agents — CurieuxExplorer · 2026-08-28