22 citations · 31 across the 7 of their papers we have counts for
4 papers · 1 filter
PeRP: Personalized Residual Policies For Congestion Mitigation Through Co-operative Advisory Systems
Aamir Hasan, Neeloy Chakraborty, Haonan Chen +3
Intelligent driving systems can be used to mitigate congestion through simple actions, thus improving many socioeconomic factors such as commute time and gas costs. However, these…
The Impact of Task Underspecification in Evaluating Deep Reinforcement Learning
Vindula Jayawardana, Catherine Tang, Sirui Li +2
Evaluations of Deep Reinforcement Learning (DRL) methods are an integral part of scientific progress of the field. Beyond designing DRL methods for general intelligence, designing…
Learning to Delegate for Large-scale Vehicle Routing
Sirui Li, Zhongxia Yan, Cathy Wu
Vehicle routing problems (VRPs) form a class of combinatorial problems with wide practical applications. While previous heuristic or learning-based works achieve decent solutions o…
Variance Reduction for Policy Gradient with Action-Dependent Factorized Baselines
Cathy Wu, Aravind Rajeswaran, Yan Duan +5
Policy gradient methods have enjoyed great success in deep reinforcement learning but suffer from high variance of gradient estimates. The high variance problem is particularly exa…