Stanford HAI questions what benchmarks measure; multilingual LLM safety paper at COLM 2026

zeeshanp_ · x · 2026-10-06

The author announces a COLM 2026 poster (Wed 11 AM, Franciscan C, #99) on multilingual safety in LLMs using multi-group IRT, co-first-authored with Max Zhang.

Context: Stanford HAI recently covered two studies asking whether benchmark scores—which influence which AI models get funded, bought, and regulated—actually measure what they claim. The work ties multilingual safety evaluation to the broader question of benchmark validity.

Original post →

More from Safety

Safety channel →