1 paper
Chi-Chang Lee, Zhang-Wei Hong, Pulkit Agrawal
In many reinforcement learning (RL) applications, augmenting the task rewards with heuristic rewards that encode human priors about how a task should be solved is crucial for achie…