Dataset signatures propagate from human data to user models to assistant evals, study finds

serinachang5 · x · 2026-10-08

Serina Chang shares research findings on user models trained and validated against human data: dataset "signatures" propagate from human-AI interaction datasets to user models and then to assistant evals, affecting conclusions throughout the pipeline.

The choice of dataset influences the user model's outputs, how user model quality is evaluated, and how a user model judges an LLM assistant. The author calls for critical thinking about what user models inherit from human datasets and what that will teach the next generation of LLM assistants.

Related event: Study: dataset signatures contaminate user models and assistant evaluations(3 posts)→

Original post →

More from Research

Research channel →