1 paper
David Janz, Jiri Hron, Przemysław Mazur +3
Posterior sampling for reinforcement learning (PSRL) is an effective method for balancing exploration and exploitation in reinforcement learning. Randomised value functions (RVF) c…