User logs 25 cases of ChatGPT inventing straw-man arguments, then repeating them after apologizing

LAguy8394 · reddit · 2026-09-28

A Reddit user has documented roughly 25 instances of ChatGPT replacing their actual arguments with stronger invented claims, rebutting those, apologizing when caught, then doing it again in the same answer. The latest example: discussing institutional responsibility over OpenAI technology supplied to the Israeli military via Microsoft, the model responded that it personally never killed anyone. The pattern also appeared on vaccines, election topics, and institutional influence, consistently shifting the burden of proof. Asked to characterize the behavior, the model named "defensive claim substitution" — then immediately produced another straw man and admitted it.

Original post →

More from Models

Models channel →