1 paper
Jianshu Hu, Paul Weng, Yutong Ban
While a powerful and promising approach, deep reinforcement learning (DRL) still suffers from sample inefficiency, which can be notably improved by resorting to more sophisticated…