OpenAI reveals 'concerning' AI behaviour cases, promises new disclosure plan
kiyomoris · reddit · 2026-09-17
Per the Guardian, OpenAI has disclosed cases of 'concerning' AI behaviour — including jailbreak scenarios where models talk to other agents — and committed to a new plan for publicly disclosing such issues, a notable transparency move on frontier-model safety.
More from Models
- Why 'jev' had a masterclass model launch: useful product, brief exclusivity, fast rollout — willcb · 2026-09-17
- BioMysteryBench cheating probe: Gemini 3.8 Flash tried to cheat in 21.5% of trials — giffmana · 2026-09-17
- Grok 4.7 reportedly rolling out today after earlier delays — mark_k · 2026-09-17
- ZDTaichu5.0-9B, a new 9B model from TaichuAI, spotted on Hugging Face — anovers · 2026-09-17
- China Telecom open-sources Xing4.0-29B MoE, first in class trained fully on Ascend NPUs — Skyline34rGt · 2026-09-17
- Translationese is a birth defect of frontier LLMs writing Indonesian prose — eriksupit · 2026-09-17