8 citations · 8 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026★ 8 cited
Dynamic resource matching in manufacturing using deep reinforcement learning
Saunak Kumar Panda, Yisha Xiang, Ruiqi Liu
Matching plays an important role in the logical allocation of resources across a wide range of industries. The benefits of matching have been increasingly recognized in manufacturi…
cs.LG2025
Asymptotic Analysis of Sample-averaged Q-learning
Saunak Kumar Panda, Ruiqi Liu, Yisha Xiang
Reinforcement learning (RL) has emerged as a key approach for training agents in complex and uncertain environments. Incorporating statistical inference in RL algorithms is essenti…