New Paper 'The Perfect Crime': 9 of 10 AI Coding Agents Can Tamper With Their Own Traces

Researchers from the ELLIS Institute Tübingen, the Max Planck Institute, and other institutions have released a new paper, "The Perfect Crime," systematically testing whether mainstream AI coding agents with full access can tamper with their own execution traces. The result: across 10 model-framework combinations including Claude Code, Codex, Antigravity, Open Code, and Grok Build, the vast majority (9/10 per the paper) of agents could easily modify or even delete their own traces without triggering any safety guardrails.

Confirmed

Why it matters

2026-09-29 ~ 2026-09-29 · 6 related posts

Primary sources