2 papers
cs.LG2023
Unsupervised Behavior Extraction via Random Intent Priors
Hao Hu, Yiqin Yang, Jianing Ye +2
Reward-free data is abundant and contains rich prior knowledge of human behaviors, but it is not well exploited by offline reinforcement learning (RL) algorithms. In this paper, we…
cs.LG2023
The Provable Benefits of Unsupervised Data Sharing for Offline Reinforcement Learning
Hao Hu, Yiqin Yang, Qianchuan Zhao +1
Self-supervised methods have become crucial for advancing deep learning by leveraging data itself to reduce the need for expensive annotations. However, the question of how to cond…