1 paper
Shilong Deng, Zetao Zheng, Hongcai He +2
A major challenge in Reinforcement Learning (RL) is the difficulty of learning an optimal policy from sparse rewards. Prior works enhance online RL with conventional Imitation Lear…