Debate: Even a Correct Eval Stops Being Right Under Massive Optimization Pressure

soumitrashukla9 · x · 2026-09-20

A debate on evals and AI safety: one side notes that with the right eval the problem is solved, but mispecification causes trouble. lugaricano pushes deeper: even an eval that is correct ex ante stops being the right one once you apply gigantic optimization pressure — a classic Goodhart-style argument in alignment discussions.

Original post →

More from AGI Musings

AGI Musings channel →