Showing cs.ROShow all
3 papers · 1 filter
cs.RO2025
MinD: Learning A Dual-System World Model for Real-Time Planning and Implicit Risk Analysis
Xiaowei Chi, Kuangzhi Ge, Jiaming Liu +9
Video Generation Models (VGMs) have become powerful backbones for Vision-Language-Action (VLA) models, leveraging large-scale pretraining for robust dynamics modeling. However, cur…
cs.RO2025
VLM-TDP: VLM-guided Trajectory-conditioned Diffusion Policy for Robust Long-Horizon Manipulation
Kefeng Huang, Tingguang Li, Yuzhen Liu +3
Diffusion policy has demonstrated promising performance in the field of robotic manipulation. However, its effectiveness has been primarily limited in short-horizon tasks, and its…
cs.RO2024
VLN-Game: Vision-Language Equilibrium Search for Zero-Shot Semantic Navigation
Bangguo Yu, Yuzhen Liu, Lei Han +3
Following human instructions to explore and search for a specified target in an unfamiliar environment is a crucial skill for mobile service robots. Most of the previous works on o…