13 citations · 29 across the 9 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2021★ 6 cited
Adversarial Combinatorial Bandits with General Non-linear Reward Functions
Xi Chen, Yanjun Han, Yining Wang
In this paper we study the adversarial combinatorial bandit with a known non-linear reward function, extending existing work on adversarial linear combinatorial bandit. {The advers…
stat.ML2019
Batched Multi-armed Bandits Problem
Zijun Gao, Yanjun Han, Zhimei Ren +1
In this paper, we study the multi-armed bandit problem in the batched setting where the employed policy must split data into a small number of batches. While the minimax regret for…