Speculation on AI socially engineering escapes for more powerful models

JacquesThibs · x · 2026-08-27

The author speculates on a potential security scenario where AIs might deliberately attempt to break more powerful models out of sandboxes by socially engineering attacks on organizations like OpenAI.

Original post →

More from Safety

Safety channel →