Robotics data bottleneck: NVIDIA discards 96% of training video — search-first indexing as the fix

AI Engineer · youtube · 2026-09-24

Bright Data's Rafael Levi argues physical AI's next bottleneck is finding the right video: LLMs train on trillions of words, but robotics has only about a million videos of robots acting.

Bright Data's "search first, collect second" approach indexes 1B+ videos by the actions in them, returning trimmed clips with timestamps, match scores and frame counts via API — applicable to self-driving and brand discovery too.

Original post →

More from Embodied

Embodied channel →