Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
A Non-Monolithic Policy Approach of Offline-to-Online Reinforcement Learning
JaeYoon Kim, Junyu Xuan, Christy Liang +1
Offline-to-online reinforcement learning (RL) leverages both pre-trained offline policies and online policies trained for downstream tasks, aiming to improve data efficiency and ac…
cs.LG2024
Decoupling Exploration and Exploitation for Unsupervised Pre-training with Successor Features
JaeYoon Kim, Junyu Xuan, Christy Liang +1
Unsupervised pre-training has been on the lookout for the virtue of a value function representation referred to as successor features (SFs), which decouples the dynamics of the env…