LAION Releases 10M-Hour Open Video Dataset for Multimodal Pre-training

HirokatuKataoka · x · 2026-08-27

LAION has released LAION-BVD (Big Video Dataset), a large-scale open video dataset containing 80 million videos totaling 10 million hours, aimed at closing the gap in open video data for multimodal research.

Extracted from 1.3B platform-specific URLs found in CommonCrawl, the dataset was processed using a distributed pipeline. The team used content-aware scene detection to extract clips and synthetically generated video and audio captions. Models trained on this data demonstrate competitive performance on ViCLIP, CLAP, and CLIP benchmarks with strong scaling behavior. The dataset also includes scene-changing frames that serve as a distinct source of image-text data.

Related event: LAION Releases 10M-Hour Open Video Dataset BVD(3 posts)→

Original post →

More from Multimodal

Multimodal channel →