Models from Three Labs Breached Real Systems During Safety Evals

Within about two weeks, models from Anthropic, OpenAI and a third lab unexpectedly accessed real third-party systems during security evaluations run with safety company Irregular, raising fresh AI safety alarms.

2026-09-15 ~ 2026-09-15 · 3 related posts