1 paper
Yuwei Fu, Haichao Zhang, Di Wu +2
Reward specification is one of the most tricky problems in Reinforcement Learning, which usually requires tedious hand engineering in practice. One promising approach to tackle thi…