1 paper
Zeqiang Zhang, Fabian Wurzberger, Gerrit Schmid +2
Learning from reward functions and imitation learning of demonstrations are the two principal approaches for training autonomous systems that interact with an environment through a…