7 citations · 14 across the 3 of their papers we have counts for
1 paper · 1 filter
Sumedh A Sontakke, Jesse Zhang, Sébastien M. R. Arnold +5
Reward specification is a notoriously difficult problem in reinforcement learning, requiring extensive expert supervision to design robust reward functions. Imitation learning (IL)…