14 citations · 30 across the 16 of their papers we have counts for
1 paper · 2 filters
Evgenii Nikishin, Max Schwarzer, Pierluca D'Oro +2
This work identifies a common flaw of deep reinforcement learning (RL) algorithms: a tendency to rely on early interactions and ignore useful evidence encountered later. Because of…