Showing 2021Show all
2 papers · 1 filter
cs.LG2021
Maximum Entropy Reinforcement Learning with Mixture Policies
Nir Baram, Guy Tennenholtz, Shie Mannor
Mixture models are an expressive hypothesis class that can approximate a rich set of policies. However, using mixture policies in the Maximum Entropy (MaxEnt) framework is not stra…
cs.LG2021
Action Redundancy in Reinforcement Learning
Nir Baram, Guy Tennenholtz, Shie Mannor
Maximum Entropy (MaxEnt) reinforcement learning is a powerful learning paradigm which seeks to maximize return under entropy regularization. However, action entropy does not necess…