2 papers
stat.ML2026
Online Statistical Inference of Constant Sample-averaged Q-Learning
Saunak Kumar Panda, Tong Li, Ruiqi Liu +1
Reinforcement learning algorithms have been widely used for decision-making tasks in various domains. However, the performance of these algorithms can be impacted by high variance…
stat.AP2026
A Statistically Reliable Optimization Framework for Bandit Experiments in Scientific Discovery
Tong Li, Travis Mandel, Goldie Phillips +4
Scientific experimentation is largely driven by statistical hypothesis testing to determine significant differences in interventions. Traditionally, experimenters allocate samples…