1 paper
Somjit Nath, Richa Verma, Abhik Ray +1
We propose a generic reward shaping approach for improving the rate of convergence in reinforcement learning (RL), called Self Improvement Based REwards, or SIBRE. The approach is…