8 citations · 12 across the 3 of their papers we have counts for
3 papers
math.OC2015★ 4 cited
Robust Policy Optimization with Baseline Guarantees
Yinlam Chow, Marek Petrik, Mohammad Ghavamzadeh
Our goal is to compute a policy that guarantees improved return over a baseline policy even when the available MDP model is inaccurate. The inaccurate model may be constructed, for…
stat.ML2012★ 8 cited
Approximate Dynamic Programming By Minimizing Distributionally Robust Bounds
Marek Petrik
Approximate dynamic programming is a popular method for solving large Markov decision processes. This paper describes a new class of approximate dynamic programming (ADP) methods-…
cs.AI2010
Global Optimization for Value Function Approximation
Marek Petrik, Shlomo Zilberstein
Existing value function approximation methods have been successfully used in many applications, but they often lack useful a priori error bounds. We propose a new approximate bilin…