3 papers
cs.RO2026
InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization
Haoxiang Ma, Junhao Cai, Xiaoxu Xu +26
Unified models for robot manipulation aim to equip one policy with both the semantic priors of pretrained VLMs and the physical dynamics learned through future prediction. In pract…
cs.RO2026
Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory
Changchuan Yang, Haoxuan Xu, Yuhang Dong +3
Conventional visuomotor imitation learning usually predicts future robot actions directly in the time domain. Such formulations often have limited physical scene awareness and weak…
cs.AI2025
ImitDiff: Transferring Foundation-Model Priors for Distraction Robust Visuomotor Policy
Yuhang Dong, Haizhou Ge, Yupei Zeng +9
Visuomotor imitation learning policies enable robots to efficiently acquire manipulation skills from visual demonstrations. However, as scene complexity and visual distractions inc…