1 paper
Yunpeng Qing, Shunyu liu, Jingyuan Cong +3
Offline reinforcement learning endeavors to leverage offline datasets to craft effective agent policy without online interaction, which imposes proper conservative constraints with…