1 citations · 1 across the 1 of their papers we have counts for
1 paper
Alessandro Abate, Yousif Almulla, James Fox +2
Training reinforcement learning (RL) agents using scalar reward signals is often infeasible when an environment has sparse and non-Markovian rewards. Moreover, handcrafting these r…