1 paper · 1 filter
Vinay Kanakeri, Shivam Bajaj, Ashwin Verma +2
It is known that reinforcement learning (RL) is data-hungry. To improve sample-efficiency of RL, it has been proposed that the learning algorithm utilize data from 'approximately s…