Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
ReferTrack: Referring Then Tracking for Embodied Visual Tracking
Hanjing Ye, Tianle Zeng, Jiazhao Zhang +6
Embodied visual tracking (EVT) requires a mobile agent to continuously follow a specific target described in natural language using only onboard vision. While recent vision-languag…
cs.RO2024
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
Kehan Chen, Dong An, Yan Huang +5
We address the task of Vision-Language Navigation in Continuous Environments (VLN-CE) under the zero-shot setting. Zero-shot VLN-CE is particularly challenging due to the absence o…