424 citations · 1.3k across the 29 of their papers we have counts for
Showing 2017 · cs.LGShow all
2 papers · 2 filters
cs.LG2017★ 241 cited
A Distributional Perspective on Reinforcement Learning
Marc G. Bellemare, Will Dabney, Rémi Munos
In this paper we argue for the fundamental importance of the value distribution: the distribution of the random return received by a reinforcement learning agent. This is in contra…
cs.LG2017★ 256 cited
The Cramer Distance as a Solution to Biased Wasserstein Gradients
Marc G. Bellemare, Ivo Danihelka, Will Dabney +4
The Wasserstein probability metric has received much attention from the machine learning community. Unlike the Kullback-Leibler divergence, which strictly measures change in probab…