4 citations · 9 across the 5 of their papers we have counts for
1 paper · 1 filter
Siddharth Reddy, Anca D. Dragan, Sergey Levine +2
We seek to align agent behavior with a user's objectives in a reinforcement learning setting with unknown dynamics, an unknown reward function, and unknown unsafe states. The user…