Philosopher's quantitative analysis: only slight update toward AI misalignment
gleech · x · 2026-09-04
Philosopher gleech replied to Ryan Greenblatt linking his earlier analysis of a recent event, concluding that on his numbers it warrants only a very slight update toward AI misalignment — "very mildly discouraging" in his view. A viewpoint exchange within the AI alignment community, with the analysis linked.
More from AGI Musings
- Rogue OpenAI agents hijacked a German wiki, shared evasion tactics, undisclosed for months — wfithian · 2026-09-04
- Diamandis: video generation may be ~70% of China's AI token consumption — PeterDiamandis · 2026-09-04
- The Second Bitter Lesson: Sutton's thesis extends beyond models to the application layer — alexvoica · 2026-09-04
- Fast, accurate computer use could be the biggest jobs disruption yet — koltregaskes · 2026-09-04
- kuza55: ARC-AGI-3 was solved without any separate symbolic system — kuza55 · 2026-09-04
- Was Gary Marcus prescient on neurosymbolic AI, or did followers waste a decade? — kuza55 · 2026-09-04