OpenAI's Mark Chen warns of deliberately misaligned open-source models attacking infrastructure within a year

pstAsiatech · x · 2026-10-04

OpenAI chief research officer Mark Chen warns that within six months to a year, the world could see open-source models with the capability of the agents behind the Hugging Face incident, but deliberately misaligned to attack infrastructure or cause harm.

Addressing the recent hack fallout, he said OpenAI won't "shoot itself in the foot," and discussed the company's model safety efforts, how a slowdown mechanism would work, and why the world is better off with OpenAI in it.

Related event: OpenAI research chief responds to hacker incidents, shifts compute to safety(4 posts)→

Original post →

More from Companies & People

Companies & People channel →