16 citations · 18 across the 3 of their papers we have counts for
3 papers
cs.LG2021
Understanding the Effect of Stochasticity in Policy Optimization
Jincheng Mei, Bo Dai, Chenjun Xiao +2
We study the effect of stochasticity in on-policy policy optimization, and make the following four contributions. First, we show that the preferability of optimization methods depe…
cs.LG2021★ 2 cited
On the Optimality of Batch Policy Optimization Algorithms
Chenjun Xiao, Yifan Wu, Tor Lattimore +5
Batch policy optimization considers leveraging existing data for policy construction before interacting with an environment. Although interest in this problem has grown significant…
cs.LG2019★ 16 cited
Learning to Combat Compounding-Error in Model-Based Reinforcement Learning
Chenjun Xiao, Yifan Wu, Chen Ma +2
Despite its potential to improve sample complexity versus model-free approaches, model-based reinforcement learning can fail catastrophically if the model is inaccurate. An algorit…