1 paper
Haozhe Liu, Mingchen Zhuge, Bing Li +4
Recent work on deep reinforcement learning (DRL) has pointed out that algorithmic information about good policies can be extracted from offline data which lack explicit information…