1 paper
Jongmin Lee, Meiqi Sun, Pieter Abbeel
In the unsupervised pre-training for reinforcement learning, the agent aims to learn a prior policy for downstream tasks without relying on task-specific reward functions. We focus…