7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.AI2017
On- and Off-Policy Monotonic Policy Improvement
Ryo Iwaki, Minoru Asada
Monotonic policy improvement and off-policy learning are two main desirable properties for reinforcement learning algorithms. In this paper, by lower bounding the performance diffe…
cs.IT2013★ 7 cited
On active information storage in input-driven systems
Oliver Obst, Joschka Boedecker, Benedikt Schmidt +1
Information theory and the framework of information dynamics have been used to provide tools to characterise complex systems. In particular, we are interested in quantifying inform…