Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Deep Policy Iteration with Integer Programming for Inventory Management
Pavithra Harsha, Ashish Jagmohan, Jayant Kalagnanam +2
We present a Reinforcement Learning (RL) based framework for optimizing long-term discounted reward problems with large combinatorial action space and state dependent constraints.…
cs.LG2024
Nonstationary Reinforcement Learning with Linear Function Approximation
Huozhi Zhou, Jinglin Chen, Lav R. Varshney +1
We consider reinforcement learning (RL) in episodic Markov decision processes (MDPs) with linear function approximation under drifting environment. Specifically, both the reward an…