38 citations · 66 across the 5 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2014★ 38 cited
Approximate Policy Iteration Schemes: A Comparison
Bruno Scherrer
We consider the infinite-horizon discounted optimal control problem formalized by Markov Decision Processes. We focus on several approximate variations of the Policy Iteration algo…
cs.AI2012★ 1 cited
Approximate Modified Policy Iteration
Bruno Scherrer, Victor Gabillon, Mohammad Ghavamzadeh +1
Modified policy iteration (MPI) is a dynamic programming (DP) algorithm that contains the two celebrated policy and value iteration methods. Despite its generality, MPI has not bee…
cs.AI2012
On the Use of Non-Stationary Policies for Infinite-Horizon Discounted Markov Decision Processes
Bruno Scherrer
We consider infinite-horizon -discounted Markov Decision Processes, for which it is known that there exists a stationary optimal policy. We consider the algorithm Value Iteratio…