1 paper
Guojian Zhan, Letian Tao, Pengcheng Wang +6
Learning expressive and efficient policy functions is a promising direction in reinforcement learning (RL). While flow-based policies have recently proven effective in modeling com…