9 citations · 9 across the 1 of their papers we have counts for
1 paper
Edoardo Cetin, Philip J. Ball, Steve Roberts +1
Off-policy reinforcement learning (RL) from pixel observations is notoriously unstable. As a result, many successful algorithms must combine different domain-specific practices and…