1 paper · 1 filter
Wei Xiao, Jiacheng Liu, Zifeng Zhuang +3
Improving the performance of pre-trained policies through online reinforcement learning (RL) is a critical yet challenging topic. Existing online RL fine-tuning methods require con…