3 papers
cs.DC2025
RollPacker: Mitigating Long-Tail Rollouts for Fast, Synchronous RL Post-Training
Wei Gao, Yuheng Zhao, Dakai An +11
Reinforcement Learning (RL) is a pivotal post-training technique for enhancing the reasoning capabilities of Large Language Models (LLMs). However, synchronous RL post-training oft…
cs.RO2025
Unified Linear Parametric Map Modeling and Perception-aware Trajectory Planning for Mobile Robotics
Hongyu Nie, Xu Liu, Zhaotong Tan +2
Autonomous navigation in mobile robots, reliant on perception and planning, faces major hurdles in large-scale, complex environments. These include heavy computational burdens for…
cs.LG2025
NAN: A Training-Free Solution to Coefficient Estimation in Model Merging
Chongjie Si, Kangtao Lv, Jingjing Jiang +6
Model merging offers a training-free alternative to multi-task learning by combining independently fine-tuned models into a unified one without access to raw data. However, existin…