OpenAI's 250-person security sprint: Codex agents wrote every patch across hundreds of systems
rohanpaul_ai · x · 2026-09-10
OpenAI published a case study and reference architecture called the "Defense Factory", describing an internal code-red sprint that mobilized 250+ people, where Codex and its cyber models wrote every patch across hundreds of systems, finding and fixing vulnerabilities humans might never have discovered.
The core argument: attackers can now run fleets of long-running agents on open-weight models, so defenders should convert security work into a continuous agent loop too — agents find vulnerabilities, validate them, and verify fixes. OpenAI shared the architecture and a practical playbook for building your own.
More from coding & agent
- AgentGrad Targets the Right Agent First: Intervention-Guided Prompt Optimization for Multi-Agent Systems — Jaewon Chu · 2026-09-10
- TUM's PlannerForge Uses LLM Agents to Automate Scenario-Based Testing of Autonomous Driving Motion Planners — TUM-AVS · 2026-09-10
- Google's free Agents Companion ebook is out for download — CodeByPoonam · 2026-09-10
- Anthropic's prompt engineering docs: when to prompt-engineer and where to start — CodeByPoonam · 2026-09-10
- Anthropic's Building Effective Agents: simple composable patterns beat complex frameworks — CodeByPoonam · 2026-09-10
- OpenAI's practical guide to building agents, plus a playbook for scaling AI use cases — CodeByPoonam · 2026-09-10