PingPong benchmark at EMNLP 2026: 6 language pairs show LLMs still struggle with code-switching

ponguru · x · 2026-09-14

The PingPong paper has been accepted to EMNLP 2026. The authors describe it as a rare, manually curated benchmark for realistic code-switched multi-party conversations, spanning 6 language combinations and 3 tasks. Their finding: current models still struggle with natural code-switching.

Related event: PingPong: Handcrafted Code-Switching Dialogue Benchmark Accepted by EMNLP 2026(2 posts)→

Original post →

More from Research

Research channel →