275 citations · 301 across the 20 of their papers we have counts for
Showing 2018 · cs.AIShow all
2 papers · 2 filters
cs.AI2018
Beyond the One Step Greedy Approach in Reinforcement Learning
Yonathan Efroni, Gal Dalal, Bruno Scherrer +1
The famous Policy Iteration algorithm alternates between policy improvement and policy evaluation. Implementations of this algorithm with several variants of the latter evaluation…
cs.AI2018★ 275 cited
Safe Exploration in Continuous Action Spaces
Gal Dalal, Krishnamurthy Dvijotham, Matej Vecerik +3
We address the problem of deploying a reinforcement learning (RL) agent on a physical system such as a datacenter cooling unit or robot, where critical constraints must never be vi…