3 papers
cs.LG2026
Dynamic resource matching in manufacturing using deep reinforcement learning
Saunak Kumar Panda, Yisha Xiang, Ruiqi Liu
Matching plays an important role in the logical allocation of resources across a wide range of industries. The benefits of matching have been increasingly recognized in manufacturi…
stat.ML2026
Online Statistical Inference of Constant Sample-averaged Q-Learning
Saunak Kumar Panda, Tong Li, Ruiqi Liu +1
Reinforcement learning algorithms have been widely used for decision-making tasks in various domains. However, the performance of these algorithms can be impacted by high variance…
cs.LG2025
Asymptotic Analysis of Sample-averaged Q-learning
Saunak Kumar Panda, Ruiqi Liu, Yisha Xiang
Reinforcement learning (RL) has emerged as a key approach for training agents in complex and uncertain environments. Incorporating statistical inference in RL algorithms is essenti…