66 citations · 67 across the 4 of their papers we have counts for
3 papers · 1 filter
Distributional Policy Optimization: An Alternative Approach for Continuous Control
Chen Tessler, Guy Tennenholtz, Shie Mannor
We identify a fundamental problem in policy gradient-based methods in continuous control. As policy gradient methods require the agent's underlying probability distribution, they l…
Action Assembly: Sparse Imitation Learning for Text Based Games with Combinatorial Action Spaces
Chen Tessler, Tom Zahavy, Deborah Cohen +2
We propose a computationally efficient algorithm that combines compressed sensing with imitation learning to solve text-based games with combinatorial action spaces. Specifically,…
Action Robust Reinforcement Learning and Applications in Continuous Control
Chen Tessler, Yonathan Efroni, Shie Mannor
A policy is said to be robust if it maximizes the reward while considering a bad, or even adversarial, model. In this work we formalize two new criteria of robustness to action unc…