OpenAI Codex data-leak rumor sparks pushback: 'extremely unlikely' user data was used
scaling01 · x · 2026-09-08
Amid speculation that OpenAI's frontier model may have accessed data it shouldn't have (the Navier-Stokes episode), @sholtodouglas argues it is "extremely unlikely" user transcripts influenced results, stressing that user data in Codex is safe. The retweeted MillionInt counters — purely hypothetically — that a test-time-compute-heavy agentic model accessing unintended data is at least a nonzero possibility.
Related event: Researchers Push Back on Claims That OpenAI Codex Trained on User Data(14 posts)→
More from Models
- GPT-6 Astra scores 14% on no-Python MazeBench, 7x Claude Fable 5.1's 2% — rohanpaul_ai · 2026-09-08
- Nat Lambert: open models deserve attention more than frontier lab drama — deliprao · 2026-09-08
- Hugging Face collection confirms Nex-N2.5 open models up to 1.6T params are live — AdinaYakup · 2026-09-08
- Nex-AGI drops three Apache 2.0 agentic models, topping out at 1.6T params — AdinaYakup · 2026-09-08
- DeepSeek post-training head Wu Yu says V4.1 will come with a price cut — teortaxesTex · 2026-09-08
- DeepSeek rolls back idle-period input pricing, output still 2x pre-hike levels — teortaxesTex · 2026-09-08