Task Scheduling as a Bandit Problem: FLD Paper Details Continuous-Time Bandit in Human Motion Learning

breadli428 · x · 2026-09-30

Responding to a question about whether the approach resembles a Boltzmann-based bandit algorithm, the author confirms that the task scheduler (teacher) making decisions from learning-progress signals is indeed a bandit problem.

An extension to the continuous-time bandit in human motion learning is exemplified in Section A.2.6 of the FLD paper.

Original post →

More from Research

Research channel →