2 papers
cs.RO2025
Restoring Noisy Demonstration for Imitation Learning With Diffusion Models
Shang-Fu Chen, Co Yong, Shao-Hua Sun
Imitation learning (IL) aims to learn a policy from expert demonstrations and has been applied to various applications. By learning from the expert policy, IL methods do not requir…
cs.LG2025
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
Ayano Hiranaka, Shang-Fu Chen, Chieh-Hsin Lai +6
Controllable generation through Stable Diffusion (SD) fine-tuning aims to improve fidelity, safety, and alignment with human guidance. Existing reinforcement learning from human fe…