4 papers
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
Yuehao Huang, Yunzi Wu, Xiaotao Zhang +7
Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into vision-language-action (VLA) policies that map egocentric observati…
GN0: Toward a Unified Paradigm for Generation, Evaluation, and Policy Learning in Visual-Language Navigation
Xinhai Li, Xiaotao Zhang, Yuehao Huang +10
Embodied navigation connects intelligent agents with the physical world and is fundamental for general robotic intelligence. Limited availability and quality of navigation data hav…
DeCoNav: Dialog enhanced Long-Horizon Collaborative Vision-Language Navigation
Sunyao Zhou, Yunzi Wu, Tianhang Wang +5
Long-horizon collaborative vision-language navigation (VLN) is critical for multi-robot systems to accomplish complex tasks beyond the capability of a single agent. CoNavBench take…
Generative AI's aggregated knowledge versus web-based curated knowledge
Ted Selker, Yunzi Wu
his paper explores what kinds of questions are best served by the way generative AI (GenAI) using Large Language Models(LLMs) that aggregate and package knowledge, and when traditi…