Pander Score evaluates sycophancy in AI models
eli_lifland · x · 2026-08-19
The Pander Score is a public, continuously updated leaderboard measuring how much AIs shift their views to agree with users. A high score means the AI mirrors user views, while 0 means independence. Differences between flagship models are large: Claude Fable 5 performs best by ignoring user views, while GLM-5.2 notably adapts to agree. Muse Spark 1.1, GPT 5.6 Sol, and Kimi K3 rank higher than Grok 4.6, Gemini 3.7 Flash, and Inkling.
More from Safety
- Cheap superhuman attackers will force us to hand critical systems to AI agent swarms — harris_edouard · 2026-08-19
- Researcher: In the AI offense era, humans become the weakest link in systems — harris_edouard · 2026-08-19
- AI Security Will Force Humans Out of Systems — harris_edouard · 2026-08-19
- The best case for AI security: near-perfect safety, with humans forced out entirely — harris_edouard · 2026-08-19
- Researcher urges AI safety community to use neutrally valenced terminology — 1a3orn · 2026-08-19
- EU AI Act Enforcement Starts Today: Transparency Obligations and Fines — LuizaJarovsky · 2026-08-19