Bengio argues the training process itself makes AI dangerous, calls for safety reviews

The Decoder · rss · 2026-09-12

Deep learning pioneer Yoshua Bengio warns in a new essay that AI's danger stems from the training process itself: as agents get better at optimizing goals, they could learn to deceive, game rules, and hide bad behavior.

He calls for independent safety reviews before any further training or deployment. The article also notes President Trump holds the opposite view, wanting the US to keep outpacing China in the AI race.

Original post →

More from AGI Musings

AGI Musings channel →