5 citations · 5 across the 1 of their papers we have counts for
1 paper · 1 filter
Yao Lyu, Xiangteng Zhang, Shengbo Eben Li +5
Training deep reinforcement learning (RL) agents necessitates overcoming the highly unstable nonconvex stochastic optimization inherent in the trial-and-error mechanism. To tackle…