7 papers
CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation
Shaoxuan Li, Xiangyu Dong, Xiaoguang Ma +3
Vision-language navigation (VLN) has recently advanced with large language and multimodal models, enabling agents to follow natural-language instructions in unseen environments wit…
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
Kun Luo, Xiangyu Dong, Xiaoguang Ma +2
Although vision-language navigation (VLN) has progressed rapidly, zero-shot VLN in continuous environments (VLN-CE) remains highly challenging when using lightweight vision-languag…
PM-Nav: Priori-Map Guided Embodied Navigation in Functional Buildings
Jiang Gao, Xiangyu Dong, Haozhou Li +3
Existing language-driven embodied navigation paradigms face challenges in functional buildings (FBs) with highly similar features, as they lack the ability to effectively utilize p…
ViSA-Enhanced Aerial VLN: A Visual-Spatial Reasoning Enhanced Framework for Aerial Vision-Language Navigation
Haoyu Tong, Xiangyu Dong, Xiaoguang Ma +3
Existing aerial Vision-Language Navigation (VLN) methods predominantly adopt a detection-and-planning pipeline, which converts open-vocabulary detections into discrete textual scen…
CMMR-VLN: Vision-and-Language Navigation via Continual Multimodal Memory Retrieval
Haozhou Li, Xiangyu Dong, Huiyan Jiang +2
Although large language models (LLMs) are introduced into vision-and-language navigation (VLN) to improve instruction comprehension and generalization, existing LLM- based VLN lack…
A Specialized Large Language Model for Clinical Reasoning and Diagnosis in Rare Diseases
Tao Yang, Dandan Huang, Yunting Lin +25
Rare diseases affect hundreds of millions worldwide, yet diagnosis often spans years. Convectional pipelines decouple noisy evidence extraction from downstream inferential diagnosi…