Agent OPFOR: Open-Source AI Agent Red Teaming Tool

grajmanu · reddit · 2026-07-08

Agent OPFOR is an open-source adversarial simulation tool for AI agents. Named after the military 'OPFOR' (Opposing Force) concept, it conducts red teaming on agents like a real adversary rather than relying on static evals. The attack surface covers multi-turn prompt injection and jailbreaks, system prompt extraction, tool abuse and BOLA/BFLA, MCP endpoint attacks (tool description injection, secret leakage, privilege escalation, SSRF), memory poisoning, goal hijacking, and EU AI Act bias testing. Its opfor hunt mode allows a command agent to autonomously plan an attack campaign given specific endpoints and targets, with support for viewing the attack tree in real time.

Related event: Agent OPFOR: Open-Source Red Teaming Tool for AI Agents(3 posts)→

Original post →

More from Safety

Safety channel →