1 paper
Pierre Le Pelletier de Woillemont, Rémi Labory, Vincent Corruble
Inverse Reinforcement Learning (IRL) learns an optimal policy, given some expert demonstrations, thus avoiding the need for the tedious process of specifying a suitable reward func…