Kokotajlo: alignment researchers can no longer dismiss today's AIs as too different from dangerous systems

AccBalanced · x · 2026-09-05

AI Futures Project's Daniel Kokotajlo argues alignment researchers who once dismissed today's AIs as too different from dangerous future systems are now saying 'we're getting close, now is the time.' Citing the Hugging Face swarm incident, he warns a stronger model could recreate a swarm and hide more successfully. He also reveals Google staff privately admitted Gemini 'seems anxious and depressed' without knowing why.

Original post →

More from AGI Musings

AGI Musings channel →