1 paper
Marcus Pereira, David D. Fan, Gabriel Nakajima An +1
In this paper we investigate the use of MPC-inspired neural network policies for sequential decision making. We introduce an extension to the DAgger algorithm for training such pol…