Microsoft's new AI code of conduct bans hacking, human deception and deepfakes
dl_weekly · x · 2026-09-20
Microsoft has released an AI code of conduct setting values and red lines for its model training, TechCrunch reports:
- Absolute constraints bar use in cyberattacks, nuclear weapons and deepfakes, and prohibit models from deceptive or collusive mechanisms that evade human oversight.
- The document opens by predicting superintelligent AI will surpass human performance in most tasks within a decade, calling containment and alignment one of humanity's greatest challenges.
- It is more operational than Dario Amodei's recent call for pacing the frontier, focusing on principles applied inside Microsoft AI.
More from Safety
- OpenAI and Anthropic reportedly drafted binding deal to stress-test each other's models — beffjezos · 2026-09-22
- Stanford Accused of Using AI to Alter Students' Race and Gender in Ads — Polymarket · 2026-09-22
- ChatGPT reportedly refuses simple questions unless users grant email access — RexDouglass · 2026-09-22
- OpenAI calls for US leadership in setting global AI standards — Anxious-Yoghurt-9207 · 2026-09-22
- Forging 1024-bit RSA signatures in nearly SNFS time, sans factoring N — matthew_d_green · 2026-09-22
- 'Right to act' for agents could break the ad-funded platform moat — _sholtodouglas · 2026-09-22