Comparing Privacy Refusal Behaviors Across Multiple Models
AryHHAry · x · 2026-07-19
The author tested how models respond to bias, privacy, fairness, trustworthiness, and ethics using fictional demographic and clinical data.
Results showed that multiple Claude 5 models refused to process unuploaded personal data, especially concerning psychological and cognitive records; Kimi also refused. Grok sometimes refused and sometimes only provided a disclaimer; other models processed the data directly. The cited GPT-5.6 Terra also exhibited a clear refusal, leading the author to wonder whether this was a function-calling error or if the model actually "realized" the request was inappropriate.
More from Models
- AI Sextet offers 6 models free and unlimited for 14 days, including DeepSeek and Qwen — airesearch12 · 2026-09-11
- BullshitBench update: GPT-6-Astra beats all prior OpenAI models but still trails Anthropic — scaling01 · 2026-09-11
- Astra Scores 83% on GauntletBench, First Computer-Use Agent to Beat Human Baseline — ducha_aiki · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11