1 paper
Chendi Qu, Jianping He, Xiaoming Duan +1
Inverse reinforcement learning (IRL) usually assumes the reward function model is pre-specified as a weighted sum of features and estimates the weighting parameters only. However,…