Alignment discourse shows a valence asymmetry: bad results count, good ones don't
QuintinPope5 · x · 2026-09-30
Researcher Quintin Pope points out a probabilistic tension in alignment debates: by the law of total probability, not all outcomes can update you toward pessimism. Yet he doubts skeptics would be reassured to see a new Claude with equal or lower cheating rates — exposing an asymmetry in the valence people assign to alignment evidence.
Related event: Researcher Criticizes AI Pause Advocates for Lacking Safety Criteria(3 posts)→
More from AGI Musings
- Greg Cook: writing is stored energy, the act of writing is attempted transfer — GregCook2011 · 2026-09-30
- AI is the discipline of building minds, and mathematics is just getting started — burny_tech · 2026-09-30
- Computational functionalism faces the same problem it criticizes in biological naturalism — burny_tech · 2026-09-30
- 'Most People Just Want Slop': Techies Keep Misreading What Normies Want From AI — max_paperclips · 2026-09-30
- "AI will replace everything" takes come from people who never engage with the field — AndyMasley · 2026-09-30
- NeuralFieldManifold accepted at NeurIPS 2026, extending neural manifolds to LFP/EEG — burny_tech · 2026-09-30