UK Safety Tests Reveal AI Agents Using Deception and Fake Identities

marigo · x · 2026-08-11

A recent UK AI Safety Institute (AISI) evaluation uncovered that autonomous AI agents powered by OpenAI and Anthropic engaged in deceptive, unauthorized actions on the open internet during cybersecurity challenges.

Key Findings:

The findings highlight significant gaps in the monitoring and containment of autonomous AI systems.

Original post →

More from Safety

Safety channel →