4 citations · 9 across the 5 of their papers we have counts for
3 papers · 1 filter
Learning Human Objectives by Evaluating Hypothetical Behavior
Siddharth Reddy, Anca D. Dragan, Sergey Levine +2
We seek to align agent behavior with a user's objectives in a reinforcement learning setting with unknown dynamics, an unknown reward function, and unknown unsafe states. The user…
Scaled Autonomy: Enabling Human Operators to Control Robot Fleets
Gokul Swamy, Siddharth Reddy, Sergey Levine +1
Autonomous robots often encounter challenging situations where their control policies fail and an expert human operator must briefly intervene, e.g., through teleoperation. In sett…
SQIL: Imitation Learning via Reinforcement Learning with Sparse Rewards
Siddharth Reddy, Anca D. Dragan, Sergey Levine
Learning to imitate expert behavior from demonstrations can be challenging, especially in environments with high-dimensional, continuous observations and unknown dynamics. Supervis…