1 paper
Hankz Hankui Zhuo, Wenfeng Feng, Yufeng Lin +2
In deep reinforcement learning, building policies of high-quality is challenging when the feature space of states is small and the training data is limited. Despite the success of…