Human attackers beat AI payment agents 74.6% of the time; deterministic layer drops rate to zero

Justgototheeffinmoon · reddit · 2026-09-21

Citing an arXiv paper (2609.22076), this post presents the first large-scale empirical proof that frontier AI payment agents are systematically defeatable via social engineering, with a 74.6% human attack success rate — and that a deterministic authorization spec layer eliminates the problem entirely. Directly relevant as agentic AI scales into financial use cases.

Original post →

More from Safety

Safety channel →