DF26 Benchmark: Humans and Detectors Both Near Random at Spotting AI Speech Videos

The new DF26 benchmark of 271 real and 2,420 synthetic public-speech videos shows both humans and top detectors perform near chance at identifying AI-generated videos, highlighting the need for robustness to distribution shift in modern generative models.

2026-09-10 ~ 2026-09-10 · 2 related posts