LVMT Sets New SOTA in Long-Term Video Segmentation While Running 10X Faster

tue-mps · hf · 2026-10-05

A new paper tackles why online video segmentation fails on long videos with long-term occlusions: temporal propagation can't adaptively select what object info to carry across time, and memory limits plus vanishing gradients prevent training on long videos.

LVMT sets a new state of the art across six benchmarks while remaining 10X faster than the prior SOTA. Code is released.

Original post →

More from Research

Research channel →