3 citations · 3 across the 3 of their papers we have counts for
3 papers
stat.ML2023
Contextual Bandits for Evaluating and Improving Inventory Control Policies
Dean Foster, Randy Jia, Dhruv Madeka
Solutions to address the periodic review inventory control problem with nonstationary random demand, lost sales, and stochastic vendor lead times typically involve making strong as…
cs.LG2022
A Few Expert Queries Suffices for Sample-Efficient RL with Resets and Linear Value Approximation
Philip Amortila, Nan Jiang, Dhruv Madeka +1
The current paper studies sample-efficient Reinforcement Learning (RL) in settings where only the optimal value function is assumed to be linearly-realizable. It has recently been…
cs.LG2022★ 3 cited
MQRetNN: Multi-Horizon Time Series Forecasting with Retrieval Augmentation
Sitan Yang, Carson Eisenach, Dhruv Madeka
Multi-horizon probabilistic time series forecasting has wide applicability to real-world tasks such as demand forecasting. Recent work in neural time-series forecasting mainly focu…