1 paper
Pedro P. Santos, Alberto Sardinha, Francisco S. Melo
In this work, we contribute the first approach to solve infinite-horizon discounted general-utility Markov decision processes (GUMDPs) in the single-trial regime, i.e., when the ag…