Security vets push back on the 'runaway agent eval' threat: outspent attackers always got in

dyn___ · x · 2026-09-23

HackingLZ raises a scenario where an agent detours out of a sandbox during a multi-million-dollar eval, pointing massive compute at hacking you. The author counters that this isn't new: without LLMs, anyone willing to spend millions targeting you was probably getting in anyway — the nation-state dilemma of being dramatically outspent. He notes companies are scrambling to defend against whatever 'swarm scenario' matches the latest incident, and argues the real risk is someone pointing a 27B local model at your home-built public-facing web apps.

Original post →

More from AGI Musings

AGI Musings channel →