2 citations · 2 across the 1 of their papers we have counts for
1 paper
Xiaoteng Ma, Xiaohang Tang, Li Xia +2
Most of reinforcement learning algorithms optimize the discounted criterion which is beneficial to accelerate the convergence and reduce the variance of estimates. Although the dis…