9 citations · 9 across the 1 of their papers we have counts for
1 paper
Haoran Xu, Xianyuan Zhan, Jianxiong Li +1
Most prior approaches to offline reinforcement learning (RL) utilize \textit{behavior regularization}, typically augmenting existing off-policy actor critic algorithms with a penal…