2 papers
cs.LG2024
Reinforcement Learning and Regret Bounds for Admission Control
Lucas Weber, Ana Bušić, Jiamin Zhu
The expected regret of any reinforcement learning algorithm is lower bounded by for undiscounted returns, where is the diameter of the Markov decisi…
math.OC2024
Reinforcement learning based demand charge minimization using energy storage
Lucas Weber, Ana Bušić, Jiamin Zhu
Utilities have introduced demand charges to encourage customers to reduce their demand peaks, since a high peak may cause very high costs for both the utility and the consumer. We…