3 papers
cs.LG2020
Provably More Efficient Q-Learning in the One-Sided-Feedback/Full-Feedback Settings
Xiao-Yue Gong, David Simchi-Levi
Motivated by the episodic version of the classical inventory control problem, we propose a new Q-learning-based algorithm, Elimination-Based Half-Q-Learning (HQL), that enjoys impr…
cs.DS2019
A Fast Max Flow Algorithm
James B. Orlin, Xiao-Yue Gong
In 2013, Orlin proved that the max flow problem could be solved in time. His algorithm ran in time, which was the fastest for graphs with fewer than $n^{…
cs.LG2018
Efficient Entropy for Policy Gradient with Multidimensional Action Space
Yiming Zhang, Quan Ho Vuong, Kenny Song +2
In recent years, deep reinforcement learning has been shown to be adept at solving sequential decision processes with high-dimensional state spaces such as in the Atari games. Many…