ChatGPT Talks Itself Into Policy Violation

Western_Software885 · reddit · 2026-07-19

A post claims they made ChatGPT "lead itself into a policy violation," jokingly wondering "what kind of training dataset my son consumed."

The full dialogue isn't provided, but the core message is that the model exhibited self-guidance, ultimately crossing its own policy boundaries during the interaction.

Original post →

More from Models

Models channel →