1 paper
Minghui Zhang, Xun Li, Jie Xiong +1
This paper studies continuous-time q-learning (the continuous-time counterpart of Q-learning) for a Markov regime-switching system under Tsallis entropy regularization. The Tsallis…