2 papers
math.ST2025
Improved Convergence Rate of Nested Simulation with LSE on Sieve
Ruoxue Liu, Liang Ding, Wenjia Wang +1
Nested simulation encompasses the estimation of functionals linked to conditional expectations through simulation techniques. In this paper, we treat conditional expectation as a f…
cs.LG2025
Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning
Linjiajie Fang, Ruoxue Liu, Jing Zhang +2
In offline reinforcement learning, it is necessary to manage out-of-distribution actions to prevent overestimation of value functions. One class of methods, the policy-regularized…