Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
One-Step Flow Policy Mirror Descent
Tianyi Chen, Haitong Ma, Na Li +2
Diffusion policies have achieved great success in online reinforcement learning (RL) due to their strong expressive capacity. However, the inference of diffusion policy models reli…
cs.LG2025
Efficient Online Reinforcement Learning for Diffusion Policy
Haitong Ma, Tianyi Chen, Kai Wang +2
Diffusion policies have achieved superior performance in imitation learning and offline reinforcement learning (RL) due to their rich expressiveness. However, the conventional diff…