3 papers
cs.RO2026
TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training
Shengbang Liu, Yueru Jia, Yuyang Yan +7
Vision-Language-Action (VLA) models have shown promising generalization in robotic manipulation, but they still struggle with contact-rich tasks, where minor contact perturbations…
cs.RO2026
Video2Act: A Dual-System Video Diffusion Policy with Robotic Spatio-Motional Modeling
Yueru Jia, Jiaming Liu, Shengbang Liu +7
Robust perception and dynamics modeling are fundamental to real-world robotic policy learning. Recent methods employ video diffusion models (VDMs) to enhance robotic policies, impr…
cs.RO2025
Decomposed Object Manipulation via Dual-Actor Policy
Bin Fan, Jian-Jian Jiang, Zhuohao Li +5
Object manipulation, which focuses on learning to perform tasks on similar parts across different types of objects, can be divided into an approaching stage and a manipulation stag…