4 papers
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
David Simchi-Levi, Zeyu Zheng, Feng Zhu
We study the optimal trade-off between expectation and tail risk for regret distribution in the stochastic multi-armed bandit model. We fully characterize the interplay among three…
Online Resource Allocation with Average Budget Constraints
Ruicheng Ao, Hongyu Chen, David Simchi-Levi +1
We consider the problem of online resource allocation with average budget constraints. At each time point the decision maker makes an irrevocable decision of whether to accept or r…
Design and Analysis of Switchback Experiments
Iavor Bojinov, David Simchi-Levi, Jinglong Zhao
Switchback experiments, where a firm sequentially exposes an experimental unit to random treatments, are among the most prevalent designs used in the technology sector, with applic…
The Competitive Ratio of Threshold Policies for Online Unit-density Knapsack Problems
Will Ma, David Simchi-Levi, Jinglong Zhao
We study a wholesale supply chain ordering problem. In this problem, the supplier has an initial stock, and faces an unpredictable stream of incoming orders, making real-time decis…