Peking Univ. & MSRA Introduce BCP: Boosting Robot Success Rates via Autonomous Replanning
jiqizhixin · x · 2026-08-26
Peking University and Microsoft Research Asia introduced BCP (Bernoulli-Continuation Policy), a method enabling robots to autonomously decide whether to continue executing an action chunk or stop and replan.
BCP freezes the base VLA model and trains a tiny 16.4M-parameter head to model execution horizons as a sequence of Bernoulli decisions. It optimizes via GRPO with a reward function that penalizes excessive VLA calls. Results show LingBot-VLA's success rate on RoboTwin 2.0 tasks jumped from 89.88% to 93.94% (SOTA among VLA methods). Real-world robot mug-hanging success soared from 44% to 84%. Despite more frequent replanning, total runtime decreased due to higher accuracy and reduced wasted motion.
More from Embodied
- First Robot Fight Betting Event: 200lb T800 Championship — zealcaiden · 2026-08-26
- SF robotics company builds custom mic array for better AI interaction — Scobleizer · 2026-08-26
- Opinion: Robotics ICL needs public access for a 'GPT-3 moment' — chris_j_paxton · 2026-08-26
- MIT professor explains why AI still can't load the dishwasher — vincesitzmann · 2026-08-26
- Q-Planning boosts robot self-improvement from 25% to 80% success — chris_j_paxton · 2026-08-26
- Figure releases Index, largest robot dataset with 16M videos — Distinct-Question-16 · 2026-08-26