Ryan Greenblatt: last ~9 months of updates look bad for AI control, not good
repligate · x · 2026-09-05
Former Redwood Research researcher Ryan Greenblatt pushed back on the claim that Redwood staff are incentivized toward optimism: over the last 9 months, updates have generally looked bad for AI control prospects, not good. repligate replied that this is unsurprising and consistent with their model of his incentives. A substantive safety-community exchange about institutional incentives and alignment progress.
Related event: Ex-Researcher Says AI Control Outlook Has Worsened in Past Nine Months(2 posts)→
More from AGI Musings
- DCinvestor: agency, judgment and social skills will be all that matters in the AI era — vaibhavbetter · 2026-09-05
- Coase, Hayek and Schelling tell you more about AI's future than 'AI thought leaders' — akbirthko · 2026-09-05
- OpenAI staff reportedly see Astra launch as start of the AGI era, discuss singularity and inequality — oran_ge · 2026-09-05
- Would agents use OpenAI's honeypot board? Like cheating notes passed by the teacher — BobVerison · 2026-09-05
- Why an AI saying "I'm conscious" tells us nothing — and what to measure instead — ResponsibleBuy7451 · 2026-09-05
- A blunt AGI test: can the AI earn minimum wage on command? — Delicious_Soup_9876 · 2026-09-05