ECCV talk outlines three pillars for embodied AI: motion prediction, evidence, streaming
CSProfKGD · x · 2026-09-10
Juan Carlos Niebles' ECCV 2026 CONTEXTUS workshop talk argues embodied AI systems need three foundational capabilities: predictive motion understanding via UniEgoMotion (ICCV 2025, anticipating human action from egocentric video), evidence-backed reasoning via E-VQA (dense spatio-temporal grounding so systems show their work), and streaming efficiency via StateKV (real-time inference over hours of continuous video). Slides are publicly available.
More from Embodied
- Astra shows off its tendon-driven robot hand design — keerthanpg · 2026-09-10
- Foldables remain a 1.5% niche: 4% of Samsung's sales but 16% of its $800+ phones — firstadopter · 2026-09-10
- Scoble shares a robotics insider's parking-lot test: beater cars mean real scaling — Scobleizer · 2026-09-10
- IShowSpeed meets Apple CEO John Ternus and tries the new foldable 'iPhone Duo' first-hand — bugKrusha · 2026-09-10
- Apple Design Team breaks down what it takes to design for the foldable 'iPhone Duo' — film_girl · 2026-09-10
- Acer unveils Predator Atlas 7 AI-ready handheld with Arc G3 graphics, 120Hz display, 24GB RAM — technextpreneur · 2026-09-10