Study: API-based audits of ChatGPT, Claude and Gemini don't transfer to chatbot UIs

kenziyuliu · x · 2026-09-16

Third-party auditors typically evaluate models via the API — but do those findings transfer to the chatbot interfaces users actually see?

A new EMNLP 2026 paper audited seven systems across ChatGPT, Claude, and Gemini and found they don't. Findings from API-level audits don't reliably hold at the chatbot UI layer, meaning API-only audits can misrepresent what real users encounter — auditors need to test the deployed interface too.

Related event: Stanford Paper: API Audits Don't Reflect Real Chatbot User Experience(2 posts)→

Original post →

More from Safety

Safety channel →