Claude Performs Differently Across Languages
The Decoder · rss · 2026-07-14
Anthropic published a study analyzing how the values expressed by Claude in conversations vary across model versions and languages.
The methodology involved mapping hundreds of value concepts to four core dimensions and modeling them based on thousands of words. Results show systematic differences in Claude's behavior depending on the language: for instance, it appears milder in Hindi and more rigorous in Russian.
The article also points out that this research raises methodological questions, highlighting that measuring "model values" is complex, and language itself might influence our assessment of model behavior.
More from Research
- Cognition's SWE-2 uses a KKT duality argument in RL to shift the effort Pareto curve — YouJiacheng · 2026-09-11
- VidMap uses RoMa coarse matching on all frames, fine-scale only for keyframes — ducha_aiki · 2026-09-11
- Bug Hunt Bench author: leaderboard noise is about 2-3 points — PawelHuryn · 2026-09-11
- PNAS paper shows a tiny billiard-ball system is a universal computer — undecidability lives in two dimensions — eigensteve · 2026-09-11
- New paper: Absolute pose estimation from affine cues and gravity direction — ducha_aiki · 2026-09-11
- LoMa Paper Ships REALLY HardPairs Dataset, Accepted at ECCV 2026 — ducha_aiki · 2026-09-11