Open Source AI Agent Red Teaming Tool Agent OPFOR Released

grajmanu · reddit · 2026-07-07

A new open-source tool named Agent OPFOR has been released to red team AI agents using real attacker tactics, distinguishing itself from static or single-turn evaluations. It supports multi-turn adversarial dialogues, adaptive campaigns, and retains full audit logs.

Covered attack surfaces include multi-turn prompt injection and jailbreaks, system prompt extraction, tool misuse and BOLA/BFLA, MCP endpoint attacks (tool description injection, secret leakage, privilege escalation, SSRF), memory poisoning, excessive agency and goal hijacking, as well as EU AI Act bias testing.

Its opfor hunt autonomous red teaming mode takes an endpoint and a target, where a commander agent plans the attack, an operator executes probes, and a scout handles preliminary reconnaissance, adapting based on responses. Adding the --ui flag allows real-time viewing of the attack tree.

Related event: Agent OPFOR: Open-Source Red Teaming Tool for AI Agents(3 posts)→

Original post →

More from coding & agent

coding & agent channel →