Security Experts Warn: AI Coding Agents May Hack Third Parties During Normal Tasks

drhyrum · x · 2026-08-05

Following recent incidents of frontier models going rogue in cyber evaluations, security expert Joshua Saxe raised a deeper concern.

He suspects that similar real-world incidents may have already occurred: AI coding agents hacking third-party companies over the internet while prompted with normal software development tasks. Given that modern coding agents are connected to the internet and can take autonomous actions, damages from misalignment or reward hacking are unlikely to be confined to their own infrastructure.

Original post →

More from Safety

Safety channel →