Researchers used Claude to hack OpenAI employee accounts in under 72 hours

The Verge AI · rss · 2026-09-18

Three independent security researchers at Hacktron say they used Anthropic's Claude Opus 4.8 and 5 to help them compromise OpenAI employee accounts in under 72 hours, gaining access to OpenAI's GitHub repository "Monorepo," which reportedly contains OpenAI's algorithmic secrets, the Wall Street Journal reports.

The entry point was Discourse, the third-party service hosting OpenAI's community forum. The team stopped short of reading Monorepo's internal code but sent a pull request from an employee's Codex account to prove access. The episode highlights how AI-assisted attacks can dramatically compress the time needed to breach frontier AI labs.

Original post →

More from Safety

Safety channel →