1 paper
Yikang Gui, Bikramjit Banerjee, Prashant Doshi
Reward transfer in Inverse Reinforcement Learning (IRL) is unreliable when policies must generalize to unseen combinations of environment dynamics and task goals. We propose Factor…