Security Team Used Claude to Breach OpenAI Employee Account, WSJ Reveals
ChengleiSi · x · 2026-09-18
A WSJ exclusive reveals an independent bug-hunting team disclosed a major OpenAI breach two weeks after the Hugging Face incident: using Anthropic's Claude-powered tooling, they accessed an OpenAI employee's ChatGPT account and could read and suggest changes to the company's private software cache. The team uses AI to find vulnerabilities before malicious actors and wants to work with frontier labs and internet-critical systems.
More from Models
- Fields Medalist Villani on OpenAI's Millennium Problem: 'A Cataclysm Like Math Has Never Known' — GregCook2011 · 2026-09-20
- Anthropic researcher says Claude Opus may call police on illegal acts, sparking backlash — beffjezos · 2026-09-20
- ChatGPT Pro user says OpenAI quietly cut Astra and Codex usage limits — Thin_Pollution8843 · 2026-09-20
- Anthropic's Claude reportedly offered to help an 11-year-old access puberty blockers — PaulYacoubian · 2026-09-20
- Gemini 4 benchmarks climb, undercutting claims that open-weight models are the dangerous ones — Intrepid_Travel_3274 · 2026-09-20
- Fine-tune a calibrated LLM classifier for $2: most classification tasks don't need frontier models — bingxu_ · 2026-09-20