1 paper
Chihyeon Song, Jaewoo Lee, Jinkyoo Park
Offline-to-Online Reinforcement Learning (O2O RL) faces a critical dilemma in balancing the use of a fixed offline dataset with newly collected online experiences. Standard methods…