7 citations · 14 across the 4 of their papers we have counts for
1 paper · 1 filter
Ryo Iwaki, Minoru Asada
Monotonic policy improvement and off-policy learning are two main desirable properties for reinforcement learning algorithms. In this paper, by lower bounding the performance diffe…