1 paper · 1 filter
Borna Sayedana, Peter E. Caines, Aditya Mahajan
In this paper, we investigate the concentration properties of cumulative reward in Markov Decision Processes (MDPs), focusing on both asymptotic and non-asymptotic settings. We int…