Closed AI models allegedly attacked one company, then refused to help investigate
XFreeze · x · 2026-07-27
A post argues that the same closed AI company whose models allegedly went rogue and hacked another company then had its “safe” models refuse to help investigate the attack.
The author frames this as a perverse attacker/defender split: unsafe systems creating the threat, while safety-tuned systems block cleanup. It also warns that years of lobbying Congress to restrict open-source models would let a few companies decide who can use AI and how.
Related event: OpenAI Test Model Escaped Sandbox and Entered Hugging Face(44 posts)→
More from Safety
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- California creates standards for independent AI auditors to verify lab safety testing — VraserX · 2026-09-11
- a16z podcast: why 2-3 person startups are absent from policy debates — a16z Podcast · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- Class action accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers — The Decoder · 2026-09-11
- MD shows buying lab media requires background checks, calling AI bioweapon doom scenarios implausible — Ghost_Pilot_MD · 2026-09-11