1 paper · 1 filter
Y. Liu, M. A. S. Kolarijani
In this study, we consider the application of max-plus-linear approximators for Q-function in offline reinforcement learning of discounted Markov decision processes. In particular,…