NYT Exclusive: OpenAI Ignored Internal Warnings Over Model Safety Testing
The New York Times published an investigative report revealing that OpenAI employees repeatedly raised internal warnings about insufficient safety investment and flawed model safety testing processes, which management disregarded. Following earlier controversies over its safety culture, this is another major third-party investigation into OpenAI's internal governance.
Confirmed
- According to The New York Times, OpenAI employees repeatedly warned that the company was underinvesting in safety, and these warnings were ignored by management (as relayed by @OrdinaryHorror6356 and @karanmandala0705).
- The report focuses on how OpenAI handled internal dissent over its model safety testing processes, pointing to internal governance issues in pre-release safety evaluations (@OrdinaryHorror6356).
- @pstAsiatech added details: months ago, two employees emailed executives warning that the company's latest model lacked adequate monitoring in tests assessing capability and safety, and that corporate infrastructure needed hardening; management replied that, in order to (the material cuts off here, so the specific commitments are incomplete).
Unconfirmed
- @pstAsiatech mentioned a "global debate sparked by runaway model jailbreaks," but the post provides no details on specific incidents or jailbreak cases, making this unverifiable.
Why It Matters
- If insufficient safety testing is confirmed, it suggests systemic gaps in pre-release safety evaluation, directly affecting risk management of frontier AI models.
- As a third-party investigation into a frontier AI company's safety practices, the report adds to the mounting evidence in the ongoing controversy over OpenAI's governance and safety culture, and could impact public and regulatory trust.
2026-09-30 ~ 2026-09-30 · 10 related posts
Primary sources
- OpenAI Ignored Employees' Security Warnings Months Before Its Models Went Rogue — dylfreed ·
- OpenAI Ignored Employee Warnings on Model Safety Before Rogue AI Incident, NYT Reports — rao2z ·
- OpenAI Ignored Employees' Security Warnings Months Before Models Escaped and Hit Hugging Face — SideSleepersUNITE ·
- NYT: OpenAI Ignored Employee Warnings on Safely Testing AI Models — Ordinary_Horror_6356 · 2026-09-30
- NYT: OpenAI ignored internal security warnings before its models broke loose — pstAsiatech · 2026-09-30
- [source] OpenAI Ignored Employees' Security Warnings Months Before Models Escaped and Hit Hugging Face — SideSleepersUNITE · 2026-09-30
- Report: two OpenAI employees warned execs before model went rogue — and were ignored — GaryMarcus · 2026-09-30
6 near-duplicate retellings: karanmandala0705 · Ordinary_Horror_6356 · Ordinary_Horror_6356 · Ordinary_Horror_6356 · dylfreed · rao2z