LAION releases 10M-hour open video dataset LAION-BVD for multimodal pre-training

thione · x · 2026-08-31

LAION released LAION-BVD, a large-scale open multimodal training dataset containing 80M videos totaling 10 million hours. Sourced from 1.3B URLs via Common Crawl, it is designed for video, audio, and image pre-training. Using content-aware scene detection, researchers extracted clips and generated synthetic captions. Models trained on this data achieve competitive performance on standard video-text and audio-text benchmarks, with consistent scaling improvements.

Original post →

More from Research

Research channel →