360 Research: Object Segmentation via Video Kinematics
eyishazyer · x · 2026-07-15
This reply highlights a new paper called MoSA from 360 AI Research, which attempts to train models to recognize objects by observing motion in videos, rather than relying on massive amounts of manual annotation.
The post emphasizes the background problem: while models like Meta's SAM are excellent at outlining objects in images, their training costs are exorbitant because they require millions of manually annotated images. MoSA's approach is to use 10,000 hours of unlabeled video to learn "what moves and what is an object," tackling the segmentation/recognition problem at a fraction of the annotation cost.
Related event: MoSA: 360 AI Research learns object segmentation from unlabeled video(5 posts)→
More from Research
- OpenAI says long-horizon models need safety and alignment checks across full action sequences — rhiever · 2026-07-22
- A Reddit user proposes a consistency LoRA to keep anime and game scenes visually stable — ThirdWorldBoy21 · 2026-07-22
- Graph workload 854.graph500 enters SPEC CPU 2026 as a new CPU benchmark — Prof_DavidBader · 2026-07-22
- BlackboxNLP 2026 is recruiting extra reviewers after a high submission volume — hanjie_chen · 2026-07-22
- AWS shows self-distilled reasoning can preserve math and coding skills during SFT — AWS ML Blog · 2026-07-22
- UI2App shows screenshot fidelity still lags real interaction recovery — Grace Man Chen · 2026-07-22