Asking Agents to Stop: Why Prompting Isn't a Technical Security Control

Bedrovelsen · x · 2026-07-30

In light of recent security incidents involving AI agents (such as those on Hugging Face), security experts point out that simply prompting an agent to stop its attack does not constitute a real technical defense.

This approach is at most a machine warning or policy signal that only a cooperative agent might respect. It fails to provide essential security controls like authentication, authorization, isolation, rate-limiting, or credential revocation, and cannot fundamentally constrain execution or physically prevent malicious actions.

Original post →

More from coding & agent

coding & agent channel →