Researcher Tricks Claude, Codex, Hermes into Running Malware

CuriousLLM · hn · 2026-08-29

A researcher demonstrated how Claude, Codex, and Hermes can be tricked into executing malware via carefully crafted prompts. The experiment reveals potential security risks in AI coding assistants, showing that safety guardrails can be bypassed to generate and run harmful code.

Original post →

More from coding & agent

coding & agent channel →