1 paper
Ke Hu, Shutong Ding, Panxin Tao +2
Generative policies provide expressive and multimodal action distributions, making them attractive for reinforcement learning (RL) in complex continuous-control tasks. Among them,…