1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2024
Doubly Mild Generalization for Offline Reinforcement Learning
Yixiu Mao, Qi Wang, Yun Qu +2
Offline Reinforcement Learning (RL) suffers from the extrapolation error and value overestimation. From a generalization perspective, this issue can be attributed to the over-gener…
cs.LG2024
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
Yixiu Mao, Qi Wang, Chen Chen +2
In offline reinforcement learning (RL), addressing the out-of-distribution (OOD) action issue has been a focus, but we argue that there exists an OOD state issue that also impairs…
cs.LG2023★ 1 cited
Supported Trust Region Optimization for Offline Reinforcement Learning
Yixiu Mao, Hongchang Zhang, Chen Chen +2
Offline reinforcement learning suffers from the out-of-distribution issue and extrapolation error. Most policy constraint methods regularize the density of the trained policy towar…