1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Jongmin Lee, Meiqi Sun, Pieter Abbeel
In the unsupervised pre-training for reinforcement learning, the agent aims to learn a prior policy for downstream tasks without relying on task-specific reward functions. We focus…