Critique of entropy decomposition in AI safety research

AdaptiveAgents · x · 2026-08-25

The author notes a cottage industry in AI safety devoted to decomposing entropy expressions and identifying terms with concepts like 'excess surprise.' While cool, they argue that someone should eventually attempt the same deconstruction with probabilities, which might yield unexpected results.

Original post →

More from Research

Research channel →