Study: dataset signatures contaminate user models and assistant evaluations
A new study finds that human-AI interaction datasets like WildChat and LMSYS carry identifiable "signatures," and these propagate through user models to contaminate downstream assistant evaluations.
2026-10-08 ~ 2026-10-08 · 3 related posts
- Study: classifiers can identify WildChat vs LMSYS chats from user messages alone — serinachang5 · 2026-10-08
- Dataset signatures propagate into user models and assistant evals, study finds — serinachang5 · 2026-10-08
- Dataset signatures propagate from human data to user models to assistant evals, study finds — serinachang5 · 2026-10-08