1 paper
Hyoungseok Kim, Jaekyeom Kim, Yeonwoo Jeong +2
Reinforcement learning algorithms struggle when the reward signal is very sparse. In these cases, naive random exploration methods essentially rely on a random walk to stumble onto…