273 citations · 304 across the 5 of their papers we have counts for
1 paper · 1 filter
Sumedh A Sontakke, Jesse Zhang, Sébastien M. R. Arnold +5
Reward specification is a notoriously difficult problem in reinforcement learning, requiring extensive expert supervision to design robust reward functions. Imitation learning (IL)…