3 papers
cs.AI2025
Landmark-Guided Knowledge for Vision-and-Language Navigation
Dongsheng Yang, Meiling Zhu, Yinfeng Yu
Vision-and-language navigation is one of the core tasks in embodied intelligence, requiring an agent to autonomously navigate in an unfamiliar environment based on natural language…
cs.AI2025
ME-BEV: Mamba-Enhanced Deep Reinforcement Learning for End-to-End Autonomous Driving with BEV-Perception
Siyi Lu, Run Liu, Dongsheng Yang +1
Autonomous driving systems face significant challenges in perceiving complex environments and making real-time decisions. Traditional modular approaches, while offering interpretab…
cs.CV2025
DOPE: Dual Object Perception-Enhancement Network for Vision-and-Language Navigation
Yinfeng Yu, Dongsheng Yang
Vision-and-Language Navigation (VLN) is a challenging task where an agent must understand language instructions and navigate unfamiliar environments using visual cues. The agent mu…