Evals are vanishing from model cards since ChatGPT went mainstream

evijit · x · 2026-09-14

The author points to their team's prior work examining which evaluations stopped being disclosed in model/system cards after ChatGPT's mainstream commercialization — and the picture is bleak. Full analysis is in their ICML paper covering 186 first-party reports and 248 third-party evaluation sources.

Related event: ICML Paper Finds AI Social Impact Evaluations Shrinking(3 posts)→

Original post →

More from Research

Research channel →