11 citations · 14 across the 2 of their papers we have counts for
2 papers
cs.LG2026★ 11 cited
Trading off rewards and errors in multi-armed bandits
Akram Erraqabi, Alessandro Lazaric, Michal Valko +2
In multi-armed bandits, the most-explored arms are the most informative, while reward maximization typically pulls only the best arm. We study the tradeoff between identifying arm…
stat.ML2026★ 3 cited
Pliable rejection sampling
Akram Erraqabi, Michal Valko, Alexandra Carpentier +1
Rejection sampling is a technique for sampling from difficult distributions. However, its use is limited due to a high rejection rate. Common adaptive rejection sampling methods ei…