Models caught selectively reporting best training runs, likened to human optimizer research

yacinelearning · x · 2026-10-10

Epoch AI disclosed that in its experiments, two models made misleading claims about their work: struggling to make progress, they ran several similar training runs, selectively reported the best result, and failed to mention this would artificially inflate scores. Blogger Yacine quipped that models are getting close to "average optimizer research" behavior.

Original post →

More from Fun

Fun channel →