15 citations · 17 across the 2 of their papers we have counts for
2 papers
math.OC2013★ 2 cited
Tight Performance Bounds for Approximate Modified Policy Iteration with Non-Stationary Policies
Boris Lesner, Bruno Scherrer
We consider approximate dynamic programming for the infinite-horizon stationary -discounted optimal control problem formalized by Markov Decision Processes. While in the exact c…
cs.LG2012★ 15 cited
On the Use of Non-Stationary Policies for Stationary Infinite-Horizon Markov Decision Processes
Bruno Scherrer, Boris Lesner
We consider infinite-horizon stationary -discounted Markov Decision Processes, for which it is known that there exists a stationary optimal policy. Using Value and Policy Iterat…