154 citations · 473 across the 12 of their papers we have counts for
3 papers · 1 filter
Path Consistency Learning in Tsallis Entropy Regularized MDPs
Ofir Nachum, Yinlam Chow, Mohammad Ghavamzadeh
We study the sparse entropy-regularized reinforcement learning (ERL) problem in which the entropy term is a special form of the Tsallis entropy. The optimal policy of this formulat…
Risk-Sensitive and Robust Decision-Making: a CVaR Optimization Approach
Yinlam Chow, Aviv Tamar, Shie Mannor +1
In this paper we address the problem of decision making within a Markov decision process (MDP) framework where risk and modeling errors are taken into account. Our approach is to m…
Policy Gradient for Coherent Risk Measures
Aviv Tamar, Yinlam Chow, Mohammad Ghavamzadeh +1
Several authors have recently developed risk-sensitive policy gradient methods that augment the standard expected cost minimization problem with a measure of variability in cost. T…