19 citations · 19 across the 1 of their papers we have counts for
1 paper
Ashish Kumar Jayant, Shalabh Bhatnagar
During initial iterations of training in most Reinforcement Learning (RL) algorithms, agents perform a significant number of random exploratory steps. In the real world, this can l…