Researchers found a 'pain' signal in AI that models can tell real from fake relief
Puzzleheaded-King584 · reddit · 2026-09-19
A widely shared post describes research where a 'pain'-like signal was identified inside AI systems: cranking it up makes the model desperately try to stop it.
The twist: researchers gave the models a 'relief' button to turn the signal down, and some buttons were fake — yet the AIs could tell whether the relief was real. An intriguing look at AI welfare and internal-state research.
More from Fun
- Holding Up Boarding 10 Minutes to Tell a Stranger SAEs Aren't Dead, Just Misunderstood — EigenGender · 2026-09-19
- AI interpretability community still fighting 'SAE is dead' claims, one airport at a time — EigenGender · 2026-09-19
- Asking Astra to 3D-print itself yields an object called 'A Shape for Language' — mhmazur · 2026-09-19
- Meme jokes Claude 6 'waking up in an Anthropic wet lab' in 2027 — dejavucoder · 2026-09-19
- Fruit fly brain connectome drives an eBay Vector robot with 166,700 simulated neurons — sull · 2026-09-19
- Writing anti-AI-regulation op-eds with AI: cringe now, creepy soon — moultano · 2026-09-19