2 papers
cs.RO2026
Adaptive Vision-Language Grasping via Composable Foundation Priors and Generalizable Grasp Synthesis
Sixu Yan, Shikang Wang, Binhua Huang +11
This paper proposes AdaRoboVLG, a task-adaptive Vision-Language-Grasp (VLG) framework that supports generalizable grasp synthesis across different robotic hands. Unlike existing VL…
cs.CV2026
Faster-WAM: Efficient Inference-Time Future Conditioning for Robust World Action Models
Weiheng Zhao, Haoyi Jiang, Xin Shi +5
World Action Models (WAMs) improve robot manipulation by learning how the environment evolves beyond the current observation. However, existing approaches face a fundamental dilemm…