Turn prediction benchmark: best systems have poor recall and high false positives

alexisgallagher · x · 2026-09-13

Citing a recent benchmark of turn prediction systems: the best existing predictors still miss many genuine turn ends (poor recall) and misidentify pauses as ends (high false positives). In the paper's corpus, turn gaps are negative for much human speech due to overlaps — and no voice system implements this accurately, the author notes.

Related event: Study: voice AI turn-taking still far from human-level(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →