collaborators

9 papers

cs.RO2026

FlashNav: Ultra-Fast Policy Training for Robot Navigation within 20 Seconds

Shanze Wang, Yiwei Qian, Xinming Zhang +6

Deep reinforcement learning has shown strong potential for robot navigation, but its practical deployment is still limited by the long wall-clock cost of policy training. This pape…

cs.RO2026

Do We Really Need Immediate Resets? Rethinking Collision Handling for Efficient Robot Navigation

Shanze Wang, Xinming Zhang, Siwei Cheng +4

Should a single collision necessarily terminate an entire navigation episode? In most deep reinforcement learning (DRL) frameworks for robot navigation, this remains the standard p…

cs.LG2026

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

Yanjun Chen, Yirong Sun, Hanlin Wang +5

Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts the result it claims to measure. This f…

cs.LG2026

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control

Jun Xue, Junze Wang, Shanze Wang +3

Scaling Maximum Entropy Reinforcement Learning (RL) to high-dimensional humanoid control remains a fundamental challenge, as the ''curse of dimensionality'' induces severe explorat…

cs.RO2026

Learning from Demonstration with Failure Awareness for Safe Robot Navigation

Xianghui Wang, Siwei Cheng, Shanze Wang +3

Learning from demonstration is widely used for robot navigation, yet it suffers from a fundamental limitation: demonstrations consist predominantly of successful behaviors and prov…

cs.RO2026

FSUNav: A Cerebrum-Cerebellum Architecture for Fast, Safe, and Universal Zero-Shot Goal-Oriented Navigation

Mingao Tan, Yiyang Li, Shanze Wang +2

Current vision-language navigation methods face substantial bottlenecks regarding heterogeneous robot compatibility, real-time performance, and navigation safety. Furthermore, they…