AI eval researcher: even humans can't detect subtle AI patterns like distributional biases

alexisjross · x · 2026-10-06

Responding to a discussion, researcher ericzelikman explains his team invests heavily in building custom evals because humans themselves are poor at detecting subtle AI patterns, especially distributional biases.

He notes that for their user models the goal is to match the actual human distribution, not to fool people — an interesting stance on how to evaluate AI systems that simulate human behavior.

Original post →

More from Models

Models channel →