3 papers
cs.LG2023
Provably Convergent Policy Optimization via Metric-aware Trust Region Methods
Jun Song, Niao He, Lijun Ding +1
Trust-region methods based on Kullback-Leibler divergence are pervasively used to stabilize policy optimization in reinforcement learning. In this paper, we exploit more flexible m…
math.OC2023
Distributionally Robust Optimal Power Flow with Uncertain Renewable Energy Output
Jia Yang, Jun Song, Chaoyue Zhao
Optimal power flow (OPF) is an important tool for Independent System Operators (ISOs) to deal with the power generation management. With the increasing penetration of renewable ene…
math.OC2023
Decision-Dependent Distributionally Robust Markov Decision Process Method in Dynamic Epidemic Control
Jun Song, William Yang, Chaoyue Zhao
In this paper, we present a Distributionally Robust Markov Decision Process (DRMDP) approach for addressing the dynamic epidemic control problem. The Susceptible-Exposed-Infectious…