1 citations · 1 across the 2 of their papers we have counts for
2 papers
stat.ML2024
Follow-the-Perturbed-Leader with Fréchet-type Tail Distributions: Optimality in Adversarial Bandits and Best-of-Both-Worlds
Jongyeong Lee, Junya Honda, Shinji Ito +1
This paper studies the optimality of the Follow-the-Perturbed-Leader (FTPL) policy in both adversarial and stochastic -armed bandits. Despite the widespread use of the Follow-th…
cs.LG2023★ 1 cited
Optimality of Thompson Sampling with Noninformative Priors for Pareto Bandits
Jongyeong Lee, Junya Honda, Chao-Kai Chiang +1
In the stochastic multi-armed bandit problem, a randomized probability matching policy called Thompson sampling (TS) has shown excellent performance in various reward models. In ad…