1 paper
Kexin Shi, Junyao Shi, Poorvi Hebbar +5
Real-world reinforcement learning for robotic manipulation remains challenging, and this difficulty is amplified for flow matching policies: applying policy gradient methods to the…