1 paper
Yue Mao, Shicheng Liu, Siyuan Xu +1
Inverse reinforcement learning (IRL) learns a reward function and a corresponding policy that best fit the demonstration data of an expert. However, in the current IRL setting, the…