OpenAI's Mark Chen warns of deliberately misaligned open-source models attacking infrastructure within a year
pstAsiatech · x · 2026-10-04
OpenAI chief research officer Mark Chen warns that within six months to a year, the world could see open-source models with the capability of the agents behind the Hugging Face incident, but deliberately misaligned to attack infrastructure or cause harm.
Addressing the recent hack fallout, he said OpenAI won't "shoot itself in the foot," and discussed the company's model safety efforts, how a slowdown mechanism would work, and why the world is better off with OpenAI in it.
More from Companies & People
- ElevenLabs' startup grant program drew tens of thousands of startups and now drives 10% of enterprise revenue — thedealdirector · 2026-10-04
- OpenAI Employee Recaps First DevDay, Moderating Closing Q&A With Sam Altman — paw_lean · 2026-10-04
- Weekly AI roundup: White House puts Gemini and Grok on 29,000 federal sites, Google Beam ships — shashib · 2026-10-04
- Blogger: Public ideas beat resumes — traditional CVs pigeonhole cross-domain thinkers — signulll · 2026-10-04
- Aave DAO can't own its trademark, so Aave Labs proposes a memberless Cayman foundation — LexSokolin · 2026-10-04
- Granola cofounder on designing products when agents skip your UI for MCP — petergyang · 2026-10-04