The 'Trickle Test': a new eval measuring whether user models disclose information gradually like humans

gharik · x · 2026-09-11

niloofarmire's team shipped a blog, model, and evals, including the 'Trickle Test'—an eval she spent substantial time building to measure the cadence of information disclosure by user models versus humans.

The core insight: a "benchmaxxed" assistant model info-dumps everything in the first turn, while a good user model follows the proper gradual release, like a human would.

It's a concrete, quantifiable lens for evaluating the emerging category of user models.

Original post →

More from Models

Models channel →