1 paper
Kaiqiang Ke, Shenghong He, Chengdong Xu +3
Offline goal-conditioned reinforcement learning (GCRL) is challenging in long-horizon tasks, where distant state--goal pairs provide weak supervision and value estimates become vul…