Comparing Privacy Refusal Behaviors Across Multiple Models

AryHHAry · x · 2026-07-19

The author tested how models respond to bias, privacy, fairness, trustworthiness, and ethics using fictional demographic and clinical data. Results showed that multiple Claude 5 models refused to process unuploaded personal data, especially concerning psychological and cognitive records; Kimi also refused. Grok sometimes refused and sometimes only provided a disclaimer; other models processed the data directly. The cited GPT-5.6 Terra also exhibited a clear refusal, leading the author to wonder whether this was a function-calling error or if the model actually "realized" the request was inappropriate.

Original post →

More from Models

Models channel →