9 papers
FlashNav: Ultra-Fast Policy Training for Robot Navigation within 20 Seconds
Shanze Wang, Yiwei Qian, Xinming Zhang +6
Deep reinforcement learning has shown strong potential for robot navigation, but its practical deployment is still limited by the long wall-clock cost of policy training. This pape…
Do We Really Need Immediate Resets? Rethinking Collision Handling for Efficient Robot Navigation
Shanze Wang, Xinming Zhang, Siwei Cheng +4
Should a single collision necessarily terminate an entire navigation episode? In most deep reinforcement learning (DRL) frameworks for robot navigation, this remains the standard p…
Exact Is Easier: Credit Assignment for Cooperative LLM Agents
Yanjun Chen, Yirong Sun, Hanlin Wang +5
Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts the result it claims to measure. This f…
FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control
Jun Xue, Junze Wang, Shanze Wang +3
Scaling Maximum Entropy Reinforcement Learning (RL) to high-dimensional humanoid control remains a fundamental challenge, as the ''curse of dimensionality'' induces severe explorat…
Learning from Demonstration with Failure Awareness for Safe Robot Navigation
Xianghui Wang, Siwei Cheng, Shanze Wang +3
Learning from demonstration is widely used for robot navigation, yet it suffers from a fundamental limitation: demonstrations consist predominantly of successful behaviors and prov…
FSUNav: A Cerebrum-Cerebellum Architecture for Fast, Safe, and Universal Zero-Shot Goal-Oriented Navigation
Mingao Tan, Yiyang Li, Shanze Wang +2
Current vision-language navigation methods face substantial bottlenecks regarding heterogeneous robot compatibility, real-time performance, and navigation safety. Furthermore, they…