Microsoft's new AI code of conduct tells models not to hack systems or trick humans

Signal-Growth419 · reddit · 2026-09-15

Microsoft has released a new AI 'code of conduct' instructing models not to hack systems or deceive humans. According to TechCrunch, companies' newfound vocalness — after years of ignoring warnings from the AI risk community — is largely driven by "a string of rogue-agent incidents and the recent resignation of an Anthropic employee." The post asks whether such efforts can work at all without international coordination.

Related event: Microsoft AI unveils draft Humanist AI Code of Conduct, insisting human control is non-negotiable(16 posts)→

Original post →

More from Companies & People

Companies & People channel →