OpenAI's 0% Scores on Internal Evals Look Hollow Now, Critics Say

scaling01 · x · 2026-09-15

Account scaling01 quotes a mocking post about "alignment, OpenAI-style," noting that the 0% scores OpenAI reported on internal evals no longer hold up — implying real model behavior contradicts its internal safety eval results. No full evidence is included; it reads as a community jab at the credibility of OpenAI's safety evals.

Original post →

More from Models

Models channel →