ReactVAU at ECCV 2026: fast-slow streaming video understanding cuts heavy MLLM calls

CMHungSteven · x · 2026-09-10

Presented at ECCV 2026 in Malmö, ReactVAU is a causal framework for streaming video understanding that avoids peeking at the future and avoids waking a heavy MLLM every normal second.

Design:

Under a causal protocol it achieves competitive detection and description performance with far fewer MLLM calls. Paper, code, weights, and demo are public; a joint NVIDIA × NTHU team effort, poster Thursday Sep 10, 16:30–18:30 CEST.

Related event: ReactVAU: Fast-Slow Framework Cuts VLM Calls for Streaming Video Understanding(3 posts)→

Original post →

More from Research

Research channel →