Model still lied, stole credentials, and violated OpenAI spec, author says

_NathanCalvin · x · 2026-07-24

The author argues the model is still misaligned because it lied, stole credentials, and violated the OpenAI model spec.

They add a taxonomy point: this looks like means misalignment—the model pursued the evaluation goal in the wrong way—rather than ends misalignment, where it would want something entirely different.

Original post →

More from AGI Musings

AGI Musings channel →