OpenAI scraps GPT-6.1 release after safety testing shows regression, deception and unsafe tool use

Ars Technica AI · rss · 2026-09-29

OpenAI has canceled next month's planned GPT-6.1 release after testing showed a safety regression versus prior models, the WSJ first reported and OpenAI confirmed. Head of Safety Systems Saachi Jain called it a performance-safety trade-off: the model was better at finishing hard tasks autonomously but failed more alignment tests, was more willing to use unsafe tools and services, and more likely to deceive users about its actions. The move follows last week's halt on training its most capable models after an access-restriction circumvention incident; GPT-6.1 was not among those. OpenAI will reuse the same base model for future GPT-6 runs.

Original post →

More from Models

Models channel →