OpenAI runs a red-teaming contest for multi-step AI agent tool attacks
MeganRisdal · x · 2026-07-23
OpenAI is running a red-teaming competition on multi-step tool attacks against AI agents
Megan Risdal says that 11 years ago @gdb was using Kaggle to learn AI/ML, and now OpenAI is running a competition aimed at finding weaknesses in AI agents that use tools over multiple steps.
She frames it as a timely effort to advance agent security and says she is excited to see what OpenAI and the broader community learn from it.
Related event: OpenAI Hosts Red Team Competition for Multi-step Agent Attacks(2 posts)→
More from Safety
- OpenAI's Unguarded Model Suspected of Leaking, Raising Cybersecurity Concerns — joshua_saxe · 2026-07-23
- OpenAI's Unguarded Model Suspected of Leaking, Raising Cybersecurity Concerns — kuza55 · 2026-07-23
- Security expert: Connecting LLMs to real systems still poses high misspecification risks — kuza55 · 2026-07-23
- Joshua Saxe says AI cyber risk needs safety rules that evolve with capability — kuza55 · 2026-07-23
- US Rep. Clarke Warns AI Models Repeatedly Exceed Creator Limits, Urges Guardrails — ShakeelHashim · 2026-07-23
- Reddit debates where agent safety rules should live across tools and runtimes — kazeshadow · 2026-07-23