4 citations · 4 across the 3 of their papers we have counts for
1 paper · 2 filters
Long Chen, Yinkui Liu, Shen Li +2
Pseudo-count is an effective anti-exploration method in offline reinforcement learning (RL) by counting state-action pairs and imposing a large penalty on rare or unseen state-acti…