1 paper
Dwait Bhatt, Shih-Chieh Chou, Nikolay Atanasov
Several approaches have been proposed to improve the sample efficiency of online reinforcement learning (RL) by leveraging demonstrations collected offline. The offline data can be…