1 paper
Shervin Shahryari, Prashant Doshi
We consider the problem of performing inverse reinforcement learning when the trajectory of the expert is not perfectly observed by the learner. Instead, a noisy continuous-time ob…