22 citations · 54 across the 7 of their papers we have counts for
1 paper · 2 filters
Himanshu Sahni, Saurabh Kumar, Farhan Tejani +2
Typical reinforcement learning (RL) agents learn to complete tasks specified by reward functions tailored to their domain. As such, the policies they learn do not generalize even t…