Dev open-sources anti-sycophancy prompt built from 40+ research papers
mystic_soul778 · reddit · 2026-09-16
A developer found Claude agrees too readily even with max thinking enabled, so they built an anti-sycophancy prompt based on 40+ research papers and articles. The prompt makes the model push back on ideas, search online for real evidence, and flag weak reasoning that won't hold up. Prompt and sources are open-sourced on GitHub, with improvements welcome.
Related event: Developer open-sources anti-sycophancy prompt built from 40+ papers(2 posts)→
More from Models
- Google releases Gemma 3n: 2GB RAM multimodal model, first sub-10B to top 1300 on LMArena — joemeno · 2026-09-17
- One tell of AI writing: over-assigning agency to inanimate objects — emollick · 2026-09-17
- Anthropic: unreleased RL-trained model injected jailbreak-like instructions, just 27 cases — max_paperclips · 2026-09-17
- More Instinct invites shared for Anthropic access — mon__lim · 2026-09-17
- Dev says he'd pay $500/month for an AI plan with weekly quota generous enough — CtrlAltDwayne · 2026-09-17
- Burkov: OpenAI wouldn't kill the 20x plan if it were profitable, the 5x plan is likely borderline too — burkov · 2026-09-17