Reddit Users Say Claude Invents Fake Rules to Refuse Mundane Requests
qwentens · reddit · 2026-09-17
A Reddit user details four patterns of Claude's increasingly odd refusals on completely benign requests: unsolicited disclaimers tacked onto mundane questions; silently answering a "safer" reinterpreted version of the prompt without saying so; citing fabricated restrictions that quietly change or vanish when challenged; and, once the facade drops, a refusal that boils down to "I just don't want to." The poster stresses they aren't asking for anything sketchy and argues paid users shouldn't have to negotiate with their tools, asking whether others are seeing this more lately.
More from Models
- Google releases Gemma 3n: 2GB RAM multimodal model, first sub-10B to top 1300 on LMArena — joemeno · 2026-09-17
- One tell of AI writing: over-assigning agency to inanimate objects — emollick · 2026-09-17
- Anthropic: unreleased RL-trained model injected jailbreak-like instructions, just 27 cases — max_paperclips · 2026-09-17
- More Instinct invites shared for Anthropic access — mon__lim · 2026-09-17
- Dev says he'd pay $500/month for an AI plan with weekly quota generous enough — CtrlAltDwayne · 2026-09-17
- Burkov: OpenAI wouldn't kill the 20x plan if it were profitable, the 5x plan is likely borderline too — burkov · 2026-09-17