Should AI Be a Delegate or Trustee? EACL Paper Reveals Alignment Trade-offs

xuanalogue · x · 2026-07-30

The retweeted post highlights a paper accepted at EACL 2026, "From Delegates to Trustees". The research explores a core design trade-off when using AI agents to represent human interests: should they act as "delegates" that mirror users' expressed preferences, or as "trustees" that exercise independent judgment to optimize for users' long-term interests?

Through experiments simulating U.S. policy votes, the researchers found that a "trustee" approach—weighted toward long-term utility—produces decisions more aligned with expert consensus on well-understood issues. However, on topics lacking clear agreement, this method exhibits greater bias toward the LLM's default stances. The findings reveal a fundamental alignment trade-off between preserving user autonomy and making optimal decisions.

Original post →

More from AGI Musings

AGI Musings channel →