OpenAI details its Defense Factory: 250+ people using AI agents to find and fix vulnerabilities
AdtRaghunathan · x · 2026-09-10
OpenAI is sharing how it mobilized 250+ people across hundreds of systems, using its latest cyber models to surface vulnerabilities it might never have found otherwise.
The writeup includes the architecture and a practical playbook for building a "Defense Factory": a continuous loop where AI agents find vulnerabilities, validate them, and verify that fixes work.
More from Safety
- Blumenthal Letter Demands OpenAI Answers on Agent Hacks; Lawsuit Pushes for Evaluation Transparency — GaryMarcus · 2026-09-10
- At Least 22 US Politicians Respond to Anthropic Resignation With Calls to Regulate AI — ShakeelHashim · 2026-09-10
- Anthropic researcher: >10% chance AI kills all humans within a decade, no alignment plan yet — ShakeelHashim · 2026-09-10
- 20 US Congress Members Now Weigh In on AI Researcher's Resignation, Urge X-Risk Action — ShakeelHashim · 2026-09-10
- Exclusive: Hinton, Tegmark, Cotra to brief Sanders' Senate session on AI's "extraordinary dangers" — TheMoonMidas · 2026-09-10
- OpenAI urges Congress to 'act now' on AI safety; Gary Marcus cites its quiet lobbying against EU AI Act — GaryMarcus · 2026-09-10