1 paper
Shuai Han, Mehdi Dastani, Shihan Wang
Improving sample efficiency is central to Reinforcement Learning (RL), especially in environments where the rewards are sparse. Some recent approaches have proposed to specify rewa…