5 papers
Best Arm Identification with LLM Judges and Limited Human
Ruicheng Ao, Hongyu Chen, Siyang Gao +2
We study fixed-confidence best-arm identification (BAI) where a cheap but potentially biased proxy (e.g., LLM judge) is available for every sample, while an expensive ground-truth…
PPI-SVRG: Unifying Prediction-Powered Inference and Variance Reduction for Semi-Supervised Optimization
Ruicheng Ao, Hongyu Chen, Haoyang Liu +2
We study semi-supervised stochastic optimization when labeled data is scarce but predictions from pre-trained models are available. PPI and SVRG both reduce variance through contro…
Learning to Price with Resource Constraints: From Full Information to Machine-Learned Prices
Ruicheng Ao, Jiashuo Jiang, David Simchi-Levi
We study the dynamic pricing problem with knapsack, addressing the challenge of balancing exploration and exploitation under resource constraints. We introduce three algorithms tai…
Prediction-Guided Active Experiments
Ruicheng Ao, Hongyu Chen, David Simchi-Levi
In this work, we introduce a new framework for active experimentation, the Prediction-Guided Active Experiment (PGAE), which leverages predictions from an existing machine learning…
Asynchronous Gradient Play in Zero-Sum Multi-agent Games
Ruicheng Ao, Shicong Cen, Yuejie Chi
Finding equilibria via gradient play in competitive multi-agent games has been attracting a growing amount of attention in recent years, with emphasis on designing efficient strate…