9 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Online Linear Programming with Batching
Haoran Xu, Peter W. Glynn, Yinyu Ye
We study Online Linear Programming (OLP) with batching. The planning horizon is cut into batches, and the decisions on customers arriving within a batch can be delayed to the e…
cs.LG2021★ 9 cited
Offline Reinforcement Learning with Soft Behavior Regularization
Haoran Xu, Xianyuan Zhan, Jianxiong Li +1
Most prior approaches to offline reinforcement learning (RL) utilize \textit{behavior regularization}, typically augmenting existing off-policy actor critic algorithms with a penal…