2 papers
cs.LG2025
Closing the Gap between TD Learning and Supervised Learning with -Conditioned Maximization
Xing Lei, Zifeng Zhuang, Shentao Yang +6
Recently, supervised learning (SL) methodology has emerged as an effective approach for offline reinforcement learning (RL) due to their simplicity, stability, and efficiency. Howe…
cs.LG2025
GCHR : Goal-Conditioned Hindsight Regularization for Sample-Efficient Reinforcement Learning
Xing Lei, Wenyan Yang, Kaiqiang Ke +4
Goal-conditioned reinforcement learning (GCRL) with sparse rewards remains a fundamental challenge in reinforcement learning. While hindsight experience replay (HER) has shown prom…