BeyondSCe grasps objects by past-event reference, hitting 77% success on occluded targets

SeoulNationalUniv · hf · 2026-10-02

Seoul National University researchers present BeyondSCe, a zero-shot robotic grasping system for event-referential requests: the target is specified by its role in a past interaction rather than by name or appearance, and may be occluded when the robot acts.

How it works:

Real-robot results with a single wrist-mounted RGB-D camera: 76% and 77% grasp success on initially visible and occluded targets (vs. 40% and 55% for the strongest baseline per condition). Across four heavily occluded scenes, success jumps from 75% to 95% while cutting mean views from 3.35 to 2.20—even beating an active-perception baseline given the target's ground-truth 3D bounding box.

Original post →

More from Embodied

Embodied channel →