1 paper · 1 filter
Kaifei Wang, Yinyu Ye, Han Zhong
Multi-armed bandit algorithms are evaluated by regret, yet comparable regret can coexist with different allocations across independent runs. We study the trade-off between worst-ca…