1 paper · 1 filter
Hwanwoo Kim, Panos Toulis, Eric Laber
Temporal difference (TD) learning is a foundational algorithm in reinforcement learning (RL). For nearly forty years, TD learning has served as a workhorse for applied RL as well a…