From the 1 of 12 linked papers with an AI index.
1 paper · 1 filter
Dong Neuck Lee, Michael R. Kosorok
Conventional off-policy reinforcement learning (RL) focuses on maximizing the expected return of scalar rewards. Distributional RL (DRL), in contrast, studies the distribution of r…