Stanford professor warns: models are increasingly opaque and brittle, hasty deployment amplifies misalignment risk
anshulkundaje · x · 2026-09-16
Stanford professor Anshul Kundaje argues the current AI paradigm is brittle and heavily dependent on training data choices, with model behavior unpredictable in unintentionally misaligned zones. As models grow more powerful and opaque, he warns these risks can't be brushed off—especially with humans deploying them hastily in flawed environments with poor safeguards.
Related event: Stanford Scholars Warn Stronger Models Are More Opaque and Fragile(2 posts)→
More from AGI Musings
- What do we call the 'it was ever thus' rhetoric used to dismiss AI concerns? — jjvincent · 2026-09-16
- AI's hardest problems need democratic deliberation — and independent experts — RishiBommasani · 2026-09-16
- Pedro Domingos: AI is anti-moat, dissolving switching costs that protect IT providers — pmddomingos · 2026-09-16
- Founder pushes back on Anthropic CEO's runaway-AI warnings on NDTV Profit — angadc · 2026-09-16
- AI Safety comms debate: punchy messaging wins short-term but erodes community epistemics — NathanpmYoung · 2026-09-16
- AI sentience debate reignites as critics call TV claims 'made-up numbers' — suchenzang · 2026-09-16