40 citations · 117 across the 21 of their papers we have counts for
1 paper · 1 filter
Siddhant Bhambri, Amrita Bhattacharjee, Durgesh Kalwar +3
Reinforcement Learning (RL) suffers from sample inefficiency in sparse reward domains, and the problem is further pronounced in case of stochastic transitions. To improve the sampl…