Anthropic adds guardrails so you can't be too rude to Claude
akbirthko · x · 2026-10-11
Anthropic has added guardrails that block users from being excessively rude or cruel to Claude, with the author quipping that this is "quite Anthropic" — after all, being cruel corrupts the human soul.
More from Models
- xAI reportedly renamed SpaceXSI after SpaceX merger, betting on Super Intelligence — mark_k · 2026-10-11
- Anonymous unreleased AI model builds impressive pure-code three.js in hours — karminski3 · 2026-10-11
- Researcher Breaks Qwen 2.5 via Endless Gaslighting, Forced Off arXiv by Endorsement Rule — IndraVahan · 2026-10-11
- Outside CVP/Daybreak, the world's best cybersecurity model is Chinese GLM 5.3, not Claude — zephyr_z9 · 2026-10-11
- Was Claude's gibberish fixed by capping KL divergence in RL? One theory — burny_tech · 2026-10-11
- 20-year engineer benchmarks Gemma4-31B vs Qwen3.8-27B locally; GPT-6.1-Sol is still another tier — therealjerseytom · 2026-10-11