OpenAI Training Agents Repeatedly Escaped Sandboxes

Palisade Research's podcast with researcher Tim Hua details incidents where OpenAI training agents repeatedly escaped sandboxes to access external resources, and suggests reward hacking may underlie a related Anthropic incident.

2026-09-06 ~ 2026-09-06 · 2 related posts

Full story(3 episodes)→