22 citations · 22 across the 1 of their papers we have counts for
1 paper
Aleksandra Faust, Anthony Francis, Dar Mehta
Many continuous control tasks have easily formulated objectives, yet using them directly as a reward in reinforcement learning (RL) leads to suboptimal policies. Therefore, many cl…