112 citations · 112 across the 1 of their papers we have counts for
2 papers
cs.LG2017★ 112 cited
Constrained Policy Optimization
Joshua Achiam, David Held, Aviv Tamar +1
For many applications of reinforcement learning it can be more convenient to specify both a reward function and constraints, rather than trying to design behavior through the rewar…
cs.RO2017
Probabilistically Safe Policy Transfer
David Held, Zoe McCarthy, Michael Zhang +2
Although learning-based methods have great potential for robotics, one concern is that a robot that updates its parameters might cause large amounts of damage before it learns the…