2 papers
cs.LG2019
On-Policy Robot Imitation Learning from a Converging Supervisor
Ashwin Balakrishna, Brijen Thananjeyan, Jonathan Lee +4
Existing on-policy imitation learning algorithms, such as DAgger, assume access to a fixed supervisor. However, there are many settings where the supervisor may evolve during polic…
cs.LG2019
Safety Augmented Value Estimation from Demonstrations (SAVED): Safe Deep Model-Based RL for Sparse Cost Robotic Tasks
Brijen Thananjeyan, Ashwin Balakrishna, Ugo Rosolia +6
Reinforcement learning (RL) for robotics is challenging due to the difficulty in hand-engineering a dense cost function, which can lead to unintended behavior, and dynamical uncert…