PRIME Benchmark: Do LLMs Rely on Social Stereotypes in Reasoning?

mdredze · x · 2026-07-03

The study introduces the PRIME benchmark to investigate whether large language models use social stereotypes as cognitive shortcuts during logical reasoning, systematically evaluating this behavior.

Original post →

More from Research

Research channel →