2 papers
cs.RO2026
FlashNav: Ultra-Fast Policy Training for Robot Navigation within 20 Seconds
Shanze Wang, Yiwei Qian, Xinming Zhang +6
Deep reinforcement learning has shown strong potential for robot navigation, but its practical deployment is still limited by the long wall-clock cost of policy training. This pape…
cs.LG2026
FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control
Jun Xue, Junze Wang, Shanze Wang +3
Scaling Maximum Entropy Reinforcement Learning (RL) to high-dimensional humanoid control remains a fundamental challenge, as the ''curse of dimensionality'' induces severe explorat…