1 paper
Xinqiang Yu, Chuanguang Yang, Chengqing Yu +3
Policy Distillation (PD) has become an effective method to improve deep reinforcement learning tasks. The core idea of PD is to distill policy knowledge from a teacher agent to a s…