2 papers
cs.RO2026
Path-conditioned Reinforcement Learning-based Local Planning for Long-Range Navigation
Mateo Haro, Julia Richter, Fan Yang +2
Long-range navigation is commonly addressed through hierarchical pipelines in which a global planner generates a path, decomposed into waypoints, and followed sequentially by a loc…
cs.LG2025
Representation Convergence: Mutual Distillation is Secretly a Form of Regularization
Zhengpeng Xie, Jiahang Cao, Changwei Wang +5
In this paper, we argue that mutual distillation between reinforcement learning policies serves as an implicit regularization, preventing them from overfitting to irrelevant featur…