VideoChat3: 4B Parameter Open Video MLLM Outperforms Larger Models

jiqizhixin · x · 2026-07-31

Researchers from Nanjing University, Shanghai AI Lab, and other institutes introduced VideoChat3, a fully open-source video multimodal large language model (MLLM) designed to efficiently process everything from short clips to hour-long streams.

Core Technical Highlights:

Performance & Availability:

Original post →

More from Multimodal

Multimodal channel →