Toby Ord Stands by RL Information-Bottleneck Thesis in Exchange with Millidge

Oxford philosopher Toby Ord, in a series of public exchanges with researcher Beren Millidge, reaffirmed his core argument about the information efficiency of reinforcement learning (RL) and offered additional explanations for why RL still works.

Confirmed

Why it matters

2026-09-23 ~ 2026-09-24 · 5 related posts

Primary sources

1 near-duplicate retellings: tobyordoxford