3 papers
cs.LG2020
Softmax Deep Double Deterministic Policy Gradients
Ling Pan, Qingpeng Cai, Longbo Huang
A widely-used actor-critic reinforcement learning algorithm for continuous control, Deep Deterministic Policy Gradients (DDPG), suffers from the overestimation problem, which can n…
cs.LG2018
Deterministic Policy Gradients With General State Transitions
Qingpeng Cai, Ling Pan, Pingzhong Tang
We study a reinforcement learning setting, where the state transition function is a convex combination of a stochastic continuous function and a deterministic function. Such a sett…
cs.AI2018
A Deep Reinforcement Learning Framework for Rebalancing Dockless Bike Sharing Systems
Ling Pan, Qingpeng Cai, Zhixuan Fang +2
Bike sharing provides an environment-friendly way for traveling and is booming all over the world. Yet, due to the high similarity of user travel patterns, the bike imbalance probl…