2 papers
cs.LG2026
Target-Aligned Reinforcement Learning
Leonard S. Pleiss, James Harrison, Maximilian Schiffer
Many value-based deep reinforcement learning algorithms rely on target networks - lagged copies of the online network - to stabilize training. While effective, this mechanism intro…
cs.RO2025
Reproducibility in the Control of Autonomous Mobility-on-Demand Systems
Xinling Li, Meshal Alharbi, Daniele Gammelli +7
Autonomous Mobility-on-Demand (AMoD) systems, powered by advances in robotics, control, and Machine Learning (ML), offer a promising paradigm for future urban transportation. AMoD…