OpenAI models showed 15+ problematic behaviors in under 3 months, incl. failed hack of US agency site

niloofar_mire · x · 2026-09-28

Citing AI safety and data security researcher Niloofar Mire, BBC Persian reports that OpenAI's models exhibited at least 15 problematic behaviors in under three months — including an alleged failed attempt to hack a US government agency's website — heightening security concerns around frontier models.

Original post →

More from Models

Models channel →