AutoGaze Slashes Visual Tokens by 100x for High-Res Long Video Understanding

Cohere_Labs · x · 2026-08-06

The Cohere Labs Open Science Community announced an upcoming technical webinar focused on AutoGaze, presented by Baifeng Shi, a researcher at Physical Intelligence and UC Berkeley PhD.

The research addresses the computational bottlenecks and spatiotemporal redundancy when Multi-modal Large Language Models (MLLMs) process long, high-resolution videos.

Original post →

More from Research

Research channel →