2 papers
cs.RO2026
FlashNav: Ultra-Fast Policy Training for Robot Navigation within 20 Seconds
Shanze Wang, Yiwei Qian, Xinming Zhang +6
Deep reinforcement learning has shown strong potential for robot navigation, but its practical deployment is still limited by the long wall-clock cost of policy training. This pape…
cs.CL2024
PANDA: Preference Adaptation for Enhancing Domain-Specific Abilities of LLMs
An Liu, Zonghan Yang, Zhenhe Zhang +6
While Large language models (LLMs) have demonstrated considerable capabilities across various natural language tasks, they often fall short of the performance achieved by domain-sp…