15 citations · 15 across the 3 of their papers we have counts for
1 paper · 1 filter
Mingyu Chen, Xuezhou Zhang
This paper initiates the study of scale-free learning in Markov Decision Processes (MDPs), where the scale of rewards/losses is unknown to the learner. We design a generic algorith…