7 citations · 25 across the 11 of their papers we have counts for
1 paper · 1 filter
Sumedh A Sontakke, Jesse Zhang, Sébastien M. R. Arnold +5
Reward specification is a notoriously difficult problem in reinforcement learning, requiring extensive expert supervision to design robust reward functions. Imitation learning (IL)…