10 citations · 21 across the 11 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Toward the Optimal Regret-Instability Trade-off in Multi-Armed Bandits
Kaifei Wang, Yinyu Ye, Han Zhong
Multi-armed bandit algorithms are evaluated by regret, yet comparable regret can coexist with different allocations across independent runs. We study the trade-off between worst-ca…
stat.ML2020
Distributionally Robust Local Non-parametric Conditional Estimation
Viet Anh Nguyen, Fan Zhang, Jose Blanchet +2
Conditional estimation given specific covariate values (i.e., local conditional estimation or functional estimation) is ubiquitously useful with applications in engineering, social…