3 papers
stat.ML2026
Variance Reduction Based Experience Replay for Policy Optimization
Hua Zheng, Wei Xie, M. Ben Feng +1
Effective reinforcement learning (RL) for complex stochastic systems requires leveraging historical data to improve sample efficiency and accelerate policy optimization. However, c…
cs.LG2026
On the Convergence of Experience Replay in Policy Optimization: Characterizing Bias, Variance, and Finite-Time Convergence
Hua Zheng, Wei Xie, M. Ben Feng
Experience replay is a core ingredient of modern deep reinforcement learning, yet its benefits in policy optimization are poorly understood beyond empirical heuristics. This paper…
q-bio.QM2024
Digital Twin Calibration for Biological System-of-Systems: Cell Culture Manufacturing Process
Fuqiang Cheng, Wei Xie, Hua Zheng
Biomanufacturing innovation relies on an efficient Design of Experiments (DoEs) to optimize processes and product quality. Traditional DoE methods, ignoring the underlying bioproce…