3 citations · 8 across the 6 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Variance Reduction Based Experience Replay for Policy Optimization
Hua Zheng, Wei Xie, M. Ben Feng +1
Effective reinforcement learning (RL) for complex stochastic systems requires leveraging historical data to improve sample efficiency and accelerate policy optimization. However, c…
stat.ML2022★ 1 cited
Variance Reduction based Experience Replay for Policy Optimization
Hua Zheng, Wei Xie, M. Ben Feng
For reinforcement learning on complex stochastic systems where many factors dynamically impact the output trajectories, it is desirable to effectively leverage the information from…