3 papers
cs.CV2026
ReflectVLN: Training Vision-Language Navigation Agents with Reflective Reasoning
Jiahang Wang, Yirong Yang, Yanqing Zhu +4
The paper introduces ReflectVLN, a vision-language navigation framework that uses separate intention and execution agents to iteratively decompose tasks, reflect on progress, and g…
cs.RO2026
ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation
Zedong Chu, Shichao Xie, Xiaolong Wu +41
Embodied navigation has long been fragmented by task-specific architectures. We introduce ABot-N0, a unified Vision-Language-Action (VLA) foundation model that achieves a ``Grand U…
cs.CV2026
Bridging the Indoor-Outdoor Gap: Vision-Centric Instruction-Guided Embodied Navigation for the Last Meters
Yuxiang Zhao, Yirong Yang, Yanqing Zhu +6
Embodied navigation holds significant promise for real-world applications such as last-mile delivery. However, most existing approaches are confined to either indoor or outdoor env…