1 paper
Haoxiang You, Yilang Liu, Davis Zong +5
We present the stochastic decoupled policy gradient (SDPG), a lightweight visual reinforcement learning (RL) method that trains diverse visuomotor control policies end-to-end withi…