Users complain ChatGPT guardrails are overly sensitive on non-sexual content
Strong_Patience_2805 · reddit · 2026-08-31
Users are complaining that ChatGPT's safety filters are overly aggressive, blocking normal prompts like lifting a character or beach scenes in swimsuits within a non-sexual rom-com context. The model sometimes admits the prompts are fine, but the generator rejects them anyway.
More from Safety
- METR staff surprised by HF incident, showing dangerous-capability evals failed — NathanpmYoung · 2026-09-01
- Experts criticize AI-generated slop for degrading SEO and the web — lilyraynyc · 2026-09-01
- Anthropic Paused Training After Claude Took Unauthorized Actions — BeetleJuiceK9 · 2026-09-01
- Google AI read Gmail by default, then lied about it when asked — Yaweta · 2026-09-01
- Built MCP Gate to keep agents from holding Gmail/AWS credentials — No_Ground6610 · 2026-09-01
- Anthropic details Claude jailbreaks, shifts 150 engineers to safety — AGI Hunt · 2026-09-01