2 papers
cs.LG2024
Matroid Semi-Bandits in Sublinear Time
Ruo-Chun Tzeng, Naoto Ohsaka, Kaito Ariu
We study the matroid semi-bandits problem, where at each round the learner plays a subset of arms from a feasible set, and the goal is to maximize the expected cumulative linea…
cs.LG2024
Best Arm Identification with Fixed Budget: A Large Deviation Perspective
Po-An Wang, Ruo-Chun Tzeng, Alexandre Proutiere
We consider the problem of identifying the best arm in stochastic Multi-Armed Bandits (MABs) using a fixed sampling budget. Characterizing the minimal instance-specific error proba…