Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Diffusion Policy through Conditional Proximal Policy Optimization
Ben Liu, Shunpeng Yang, Hua Chen
Reinforcement learning (RL) has been extensively employed in a wide range of decision-making problems, such as games and robotics. Recently, diffusion policies have shown strong po…
cs.LG2026
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
Shunpeng Yang, Ben Liu, Hua Chen
Among on-policy reinforcement learning algorithms, Proximal Policy Optimization (PPO) demonstrates is widely favored for its simplicity, numerical stability, and strong empirical p…