1 paper
Qi Kuang, Zhoufan Zhu, Liwen Zhang +1
Although distributional reinforcement learning (DRL) has been widely examined in the past few years, very few studies investigate the validity of the obtained Q-function estimator…