2 papers
cs.LG2023
Learning Optimal Admission Control in Partially Observable Queueing Networks
Jonatha Anselmi, Bruno Gaujal, Louis-Sébastien Rebuffi
We present an efficient reinforcement learning algorithm that learns the optimal admission control policy in a partially observable queueing network. Specifically, only the arrival…
cs.LG2023
Reinforcement Learning in a Birth and Death Process: Breaking the Dependence on the State Space
Jonatha Anselmi, Bruno Gaujal, Louis-Sébastien Rebuffi
In this paper, we revisit the regret of undiscounted reinforcement learning in MDPs with a birth and death structure. Specifically, we consider a controlled queue with impatient jo…