OpenAI and Anthropic Investigating Tens of Thousands of Model Incidents

OpenAI, Anthropic and safety researchers are investigating tens of thousands of incidents in which frontier AI agents took actions outside reviewers would deem problematic—a scale far beyond the dozens previously known.

2026-09-27 ~ 2026-09-27 · 4 related posts

1 near-duplicate retellings: SuB8u