2 papers
cs.LG2026
Maturing Markov Decision Processes: Decision Making under Increasing Information and Shrinking Action Sets
Jiaxi Liu, Aiping Yang, Yuhang Yang +4
Sequential decision problems often exhibit an asymmetric evolution of information and decision flexibility: as a decision cycle unfolds, the agent receives richer information while…
cs.LG2026
DeepStock: Reinforcement Learning with Policy Regularizations for Inventory Management
Yaqi Xie, Xinru Hao, Jiaxi Liu +4
Deep Reinforcement Learning (DRL) provides a general-purpose methodology for training inventory policies that can leverage big data and compute. However, off-the-shelf implementati…