Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Preference is More Than Comparisons: Rethinking Dueling Bandits with Augmented Human Feedback
Shengbo Wang, Hong Sun, Ke Li
Interactive preference elicitation (IPE) aims to substantially reduce human effort while acquiring human preferences in wide personalization systems. Dueling bandit (DB) algorithms…
cs.LG2023
Constrained Bayesian Optimization Under Partial Observations: Balanced Improvements and Provable Convergence
Shengbo Wang, Ke Li
The partially observable constrained optimization problems (POCOPs) impede data-driven optimization techniques since an infeasible solution of POCOPs can provide little information…