1 paper
Daniel Koutas, Daniel Hettegger, Kostas G. Papakonstantinou +1
We present a novel method for Deep Reinforcement Learning (DRL), incorporating the convex property of the value function over the belief space in Partially Observable Markov Decisi…