Princeton's PAST-Bench Tests If Personal Agents Actually Improve From Accumulated Experience

princetonu · hf · 2026-08-05

A Princeton team introduced PAST-Bench, a benchmark designed to evaluate whether personal AI agents can translate retained experience (preferences, histories, skills) into better future performance.

Original post →

More from coding & agent

coding & agent channel →