29 citations · 29 across the 1 of their papers we have counts for
1 paper
Borislav Mavrin, Shangtong Zhang, Hengshuai Yao +3
In distributional reinforcement learning (RL), the estimated distribution of value function models both the parametric and intrinsic uncertainties. We propose a novel and efficient…