Debating FDT's decade of stalled progress: ex ante optimality as an alternative frame

jessi_cata · x · 2026-10-09

In a debate with @MatriceJacobine, jessicata argues that MIRI-ish researchers have been applying the same "FDT! TDT! logical counterfactuals!" frame for over a decade without success, and suggests working with different concepts for thinking about ex ante optimality and stability under self-modification.

Pressed further, she adds that she isn't convinced FDT is well-posed, though discussing ex ante optimality in various situations (e.g. imperfect recall) is meaningful, and extension to Newcomblike problems may be possible. A substantive clash over whether decision-theory foundations are a dead end for AI alignment research.

Original post →

More from AGI Musings

AGI Musings channel →