Opus 5.5 safeguards flag harmless queries about currency, hotel points and noodles
ElvisGrizzly · reddit · 2026-09-24
A Reddit user catalogued a series of completely benign questions that triggered Opus 5.5's "safeguards flagged this message" error in a single day: comparing $897.45 vs 150,000 Hilton points, asking about Amex hotel transfer partners, weighing $319 cash vs 45,000 Amex points for a hotel stay, asking how to best eat dry chili pan mee in Kuala Lumpur, and converting 250,000 yen to dollars.
The poster quips that Claude is "prejudiced against maximizing value on hotel points, Japanese money and noodles" — a vivid example of over-aggressive safety guardrails misfiring on everyday queries.
More from Models
- Dhravya Shah Details How typesafeAI's Jev Improves AI Memory and Context Engineering — blaizedsouza · 2026-09-25
- Community Speculates DeepSeek V4 Coming Soon After Holiday Post From Liang Wenfeng — teortaxesTex · 2026-09-25
- NaceAI launches Drex, a sub-6B decision model that tops the public Decision Index at 51.73 — ordax · 2026-09-25
- Model audit showdown: Astra dominates, Opus and Fable close, Grok 4.7 and GPT-6 Sol lag far behind — ivan_bezdomny · 2026-09-25
- Uncensored local model Bonzai 2 27B tops benchmarks, runs on 12GB VRAM — alexcovo_eth · 2026-09-25
- Same prompt, Opus 5.5 one-shot video generation put to a public retest with different tools — drrickio · 2026-09-25