103 citations · 210 across the 13 of their papers we have counts for
1 paper · 1 filter
Weijun Hong, Menghui Zhu, Minghuan Liu +4
Exploration is crucial for training the optimal reinforcement learning (RL) policy, where the key is to discriminate whether a state visiting is novel. Most previous work focuses o…