1 paper
Moritz A. Zanger, Wendelin Böhmer, Matthijs T. J. Spaan
In contrast to classical reinforcement learning (RL), distributional RL algorithms aim to learn the distribution of returns rather than their expected value. Since the nature of th…