4 papers
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
Shunpeng Yang, Ben Liu, Hua Chen
Among on-policy reinforcement learning algorithms, Proximal Policy Optimization (PPO) demonstrates is widely favored for its simplicity, numerical stability, and strong empirical p…
See Once, Then Act: Vision-Language-Action Model with Task Learning from One-Shot Video Demonstrations
Guangyan Chen, Meiling Wang, Qi Shao +10
Developing robust and general-purpose manipulation policies represents a fundamental objective in robotics research. While Vision-Language-Action (VLA) models have demonstrated pro…
LIPM-Guided Reinforcement Learning for Stable and Perceptive Locomotion in Bipedal Robots
Haokai Su, Haoxiang Luo, Shunpeng Yang +3
Achieving stable and robust perceptive locomotion for bipedal robots in unstructured outdoor environments remains a critical challenge due to complex terrain geometry and susceptib…
Multi-Loco: Unifying Multi-Embodiment Legged Locomotion via Reinforcement Learning Augmented Diffusion
Shunpeng Yang, Zhen Fu, Zhefeng Cao +4
Generalizing locomotion policies across diverse legged robots with varying morphologies is a key challenge due to differences in observation/action dimensions and system dynamics.…