Moritz Hardt's New Book on ML Benchmarking Science Lands Amid Eval Audit Drama

JJitsev · x · 2026-09-20

Researcher JJitsev highlights the recent debate on how to audit ML evals, what counts as eval flaws, and how audits themselves can be flawed — and recommends Moritz Hardt's new book The Emerging Science of Machine Learning Benchmarks (Princeton University Press, hardcover October 6, 2026).

Original post →

More from Models

Models channel →