1 paper · 1 filter
Guopeng Li, Yiyang Duan, Yiru Jiao +1
Contrastive reinforcement learning (CRL) scales effectively in goal-conditioned tasks by casting policy learning into a self-supervised contrastive objective. However, in a failure…