Stanford HAI warns that averaging expert safety scores can erase good chatbot advice

StanfordHAI · x · 2026-07-25

Stanford HAI says AI developers often rely on mental health experts to judge whether chatbot answers are safe — but averaging expert scores can produce advice nobody would actually call good.

Original post →

More from Research

Research channel →