Where Are All the Prompt Injection Damages?

joshua_saxe · x · 2026-08-23

After evangelizing about prompt injection risks, the author observed minimal real-world damages compared to overall cyber losses. Analysis suggests universal jailbreaks may have become harder, especially for Anthropic and GPT-5.6 Sol models. However, goal hijacking remains practical. The article warns that while fixing global security tech debt is expensive, agents could make exploiting these vulnerabilities significantly cheaper for attackers.

Original post →

More from Safety

Safety channel →