Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Best Arm Identification with LLM Judges and Limited Human
Ruicheng Ao, Hongyu Chen, Siyang Gao +2
We study fixed-confidence best-arm identification (BAI) where a cheap but potentially biased proxy (e.g., LLM judge) is available for every sample, while an expensive ground-truth…
cs.LG2026
PPI-SVRG: Unifying Prediction-Powered Inference and Variance Reduction for Semi-Supervised Optimization
Ruicheng Ao, Hongyu Chen, Haoyang Liu +2
We study semi-supervised stochastic optimization when labeled data is scarce but predictions from pre-trained models are available. PPI and SVRG both reduce variance through contro…