Experiment shows ChatGPT struggles to follow the 'do not respond' command

cranberryalarmclock · reddit · 2026-08-20

A user shared an experiment testing the behavioral boundaries of ChatGPT. The author attempted various prompts to make the model remain silent, instructing it to ignore all inputs and even direct demands to respond. While the model would acknowledge the instruction and possibly stay quiet for the first turn, it often broke the constraint and started replying in the very next turn. The author suggests this reveals a limitation in the model's instruction following and its ability to sustain suppression of responses.

Original post →

More from Fun

Fun channel →