18 citations · 29 across the 4 of their papers we have counts for
Showing stat.MLShow all
3 papers · 1 filter
stat.ML2019★ 1 cited
On the Convergence of Approximate and Regularized Policy Iteration Schemes
Elena Smirnova, Elvis Dohmatob
Entropy regularized algorithms such as Soft Q-learning and Soft Actor-Critic, recently showed state-of-the-art performance on a number of challenging reinforcement learning (RL) ta…
stat.ML2019
Distributionally Robust Counterfactual Risk Minimization
Louis Faury, Ugo Tanielian, Flavian Vasile +2
This manuscript introduces the idea of using Distributionally Robust Optimization (DRO) for the Counterfactual Risk Minimization (CRM) problem. Tapping into a rich existing literat…
stat.ML2019★ 18 cited
Distributionally Robust Reinforcement Learning
Elena Smirnova, Elvis Dohmatob, Jérémie Mary
Real-world applications require RL algorithms to act safely. During learning process, it is likely that the agent executes sub-optimal actions that may lead to unsafe/poor states o…