1 paper
Nicole Bäuerle, Marcin Pitera, Łukasz Stettner
We study discrete-time Markov Decision Processes (MDPs) on finite state-action spaces and analyze the stability of optimal policies and value functions in the long-run discounted r…