2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2023
COPlanner: Plan to Roll Out Conservatively but to Explore Optimistically for Model-Based RL
Xiyao Wang, Ruijie Zheng, Yanchao Sun +4
Dyna-style model-based reinforcement learning contains two phases: model rollouts to generate sample for policy learning and real environment exploration using current policy for d…
cs.LG2022★ 2 cited
Live in the Moment: Learning Dynamics Model Adapted to Evolving Policy
Xiyao Wang, Wichayaporn Wongkamjan, Furong Huang
Model-based reinforcement learning (RL) often achieves higher sample efficiency in practice than model-free RL by learning a dynamics model to generate samples for policy learning.…