9 papers · 1 filter
Retrieve in Time, Correct in Frequency
Yuze Fan, Yue Cao, Pengjie Gao +7
Frozen vision-language-action (VLA) policies generate temporally extended action chunks, but long-horizon manipulation remains vulnerable to accumulated execution error and visual…
RPG: Robust Policy Gating for Smooth Multi-Skill Transitions in Humanoid Fighting
Yucheng Xin, Jiacheng Bao, Yubo Dong +5
Humanoid robots have demonstrated impressive motor skills in a wide range of tasks, yet whole-body control for humanlike long-time, dynamic fighting remains particularly challengin…
Learn Weightlessness: Imitate Non-Self-Stabilizing Motions on Humanoid Robot
Yucheng Xin, Jiacheng Bao, Haoran Yang +6
The integration of imitation and reinforcement learning has enabled remarkable advances in humanoid whole-body control, facilitating diverse human-like behaviors. However, research…
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization
Yihao Wu, He Zhang, Junbo Tan +2
Post-training Vision-Language-Action (VLA) models into policies that can be reliably deployed on real robots remains a major bottleneck. SFT and DAgger exploit failure signals only…
Choose What to Manipulate: Revealing Data Scaling Laws in Bounding-Box Guided Policies for Semantic Manipulation
Yihao Wu, Jinming Ma, Junbo Tan +5
Diffusion-based policies generalize poorly in semantic manipulation, a key obstacle to real-world deployment, because text-only instructions cannot reliably steer the policy toward…
A Hybrid Force-Position Strategy for Shape Control of Deformable Linear Objects With Graph Attention Networks
Yanzhao Yu, Haotian Yang, Junbo Tan +1
Manipulating deformable linear objects (DLOs) such as wires and cables is crucial in various applications like electronics assembly and medical surgeries. However, it faces challen…