3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2022
Linear Reinforcement Learning with Ball Structure Action Space
Zeyu Jia, Randy Jia, Dhruv Madeka +1
We study the problem of Reinforcement Learning (RL) with linear function approximation, i.e. assuming the optimal action-value function is linear in a known -dimensional feature…
cs.LG2019★ 3 cited
Learning in structured MDPs with convex cost functions: Improved regret bounds for inventory management
Shipra Agrawal, Randy Jia
We consider a stochastic inventory control problem under censored demands, lost sales, and positive lead times. This is a fundamental problem in inventory management, with signific…