1 paper
Changmin Yu, David Mguni, Dong Li +3
Efficient reinforcement learning (RL) involves a trade-off between "exploitative" actions that maximise expected reward and "explorative'" ones that sample unvisited states. To enc…