11 citations · 21 across the 5 of their papers we have counts for
1 paper · 1 filter
Himanshu Sahni, Saurabh Kumar, Farhan Tejani +2
Typical reinforcement learning (RL) agents learn to complete tasks specified by reward functions tailored to their domain. As such, the policies they learn do not generalize even t…