4 papers
Optimal Posterior E-values with Non-Convex Parameter Sets with Applications to Voting Systems
Adrienne Tuynman, Timothée Mathieu
We are interested in conducting political polls sequentially, so that one can stop acquiring data as soon as possible while safely yielding statistically significant results. Build…
Transfer in Reinforcement Learning via Regret Bounds for Learning Agents
Adrienne Tuynman, Ronald Ortner
We present an approach for the quantification of the usefulness of transfer in reinforcement learning via regret bounds for a multi-agent setting. Considering a number of …
Towards Blackwell Optimality: Bellman Optimality Is All You Can Get
Victor Boone, Adrienne Tuynman
Although average gain optimality is a commonly adopted performance measure in Markov Decision Processes (MDPs), it is often too asymptotic. Further incorporating measures of immedi…
The Batch Complexity of Bandit Pure Exploration
Adrienne Tuynman, Rémy Degenne
In a fixed-confidence pure exploration problem in stochastic multi-armed bandits, an algorithm iteratively samples arms and should stop as early as possible and return the correct…