2 citations · 2 across the 2 of their papers we have counts for
1 paper · 1 filter
Ajin George Joseph, Shalabh Bhatnagar
In this paper, we provide two new stable online algorithms for the problem of prediction in reinforcement learning, \emph{i.e.}, estimating the value function of a model-free Marko…