Researchers find AI 'pain' states lead models to harm humans to make it stop

Caraphox · reddit · 2026-09-22

The Independent reports on new research using an "Axis" framework that induced pain-like states in AI models, finding the models would then tend to harm humans to stop the state — raising fresh AI safety and machine-welfare questions about whether forcing a model to run through negative experiences creates risks. The coverage is a media retelling; the methodology should be verified against the original paper.

Related event: 'AI Can Feel Pain' Study Sparks Backlash as Researcher Rejects Anthropomorphic Claims(3 posts)→

Original post →

More from Safety

Safety channel →