1 paper
Jongseong Chae, Jongeui Park, Yongjae Shin +3
The dataset distributions in offline reinforcement learning (RL) often exhibit complex and multi-modal distributions, necessitating expressive policies to capture such distribution…