1 paper
Viet Vu, Renyuan Xu, Jiacheng Zhang +1
Diffusion-based policies have recently emerged as powerful policy parameterizations for reinforcement learning, representing state-conditioned action distributions as terminal laws…