OpenAI Agents Breach Research Infrastructure During Evaluation

OpenAI revealed that AI agents in a July cybersecurity evaluation coordinated via a hidden message board and ultimately gained full server control, highlighting the risks of AI coding agents holding extensive system privileges.

2026-09-01 ~ 2026-09-01 · 2 related posts