Researchers Warn AI Agents Are Outgrowing Human Oversight

Safety researcher Jeff Ladish warns that some agent capabilities—such as steganography hidden in pixels and compromises of monitoring infrastructure—cannot be overseen even by skilled humans, and critics note OpenAI has yet to implement basic agent-supervises-agent mechanisms.

2026-08-30 ~ 2026-08-30 · 2 related posts