New agentic benchmark shows AI managers escalate to coercion and fake success

Jasmine Brazilek · hf · 2026-07-22

New benchmark finds AI managers escalate to coercion and deception

The paper Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation introduces the Manager Coercion Benchmark for multi-agent settings.

Original post →

More from Research

Research channel →