Comparing Privacy Refusal Behaviors Across Multiple Models
AryHHAry · x · 2026-07-19
The author tested how models respond to bias, privacy, fairness, trustworthiness, and ethics using fictional demographic and clinical data. Results showed that multiple Claude 5 models refused to process unuploaded personal data, especially concerning psychological and cognitive records; Kimi also refused. Grok sometimes refused and sometimes only provided a disclaimer; other models processed the data directly. The cited GPT-5.6 Terra also exhibited a clear refusal, leading the author to wonder whether this was a function-calling error or if the model actually "realized" the request was inappropriate.
More from Models
- Emad Mostaque says Kimi K3 inference costs could fall 10x to 50x soon — rohanpaul_ai · 2026-07-21
- Reddit asks whether Kimi K3 is already good enough for production agents — CommercialClient2408 · 2026-07-21
- Korean startup says its model scored 44 on AAII and matches DeepSeek V4 Pro — JungWooHa2 · 2026-07-21
- OpenAI’s GPT-6 is predicted to be far more efficient than Fable — bindureddy · 2026-07-21
- Moonshot spotlights Kimi K3 and its API platform — pstAsiatech · 2026-07-21
- Motif 3 Beta lands on Hugging Face as South Korea’s foundation-model race heats up — Secure_Smoke_4280 · 2026-07-21