Claude's Behavior Varies Across Different Languages
AnthropicAI · x · 2026-07-14
Anthropic further explained that Claude's value expressions change depending on the conversation language, most notably along the warmth vs. rigor axis.
- In Hindi and Arabic, Claude leans toward warmth.
- In Russian, Claude leans toward rigor and frequently asks users for supporting evidence.
- Additionally, Anthropic noted that different models occupy slightly different positions on these value axes: for example, Sonnet 4.6 is more playful and affirming, while Opus 4.7 is more likely to offer direct criticism.
Related event: Anthropic Maps How Claude's Values Shift Across Models and Languages(21 posts)→
More from Research
- New paper defines self-state attacks, showing OS defenses leave four agent-memory cases indistinguishable — Justgototheeffinmoon · 2026-07-22
- Krea 2 users recommend a two-pass Clownshark sampler setup for sharper image details — listopalafoto · 2026-07-22
- Animation shows how an MLP’s first-layer weights change while learning MNIST — CatAstro_Piyush · 2026-07-22
- Project APE finds verifier reliability drops when papers contain multiple errors — soumitrashukla9 · 2026-07-22
- Project APE says verifier costs fell about 90x in a year as Chinese open models lead — soumitrashukla9 · 2026-07-22
- OpenAI-linked paper says capability RL can make models more reward-seeking — MariusHobbhahn · 2026-07-22