activity
20232026
collaborators
Showing cs.LGShow all

6 papers · 1 filter

cs.LG2026

Variance-sensitive Thompson sampling for generalised linear bandits, revisited

Tom Perneczky, Marc Abeille, David Janz

We prove a variance-sensitive regret bound for Thompson sampling in stochastic generalised linear bandits. The argument assumes a warm-up, after which the regret is controlled thro…

cs.LG2026

Sharp analysis of linear ensemble sampling

David Janz, Arya Akhavan, Csaba Szepesvári

We analyse linear ensemble sampling (ES) with standard Gaussian perturbations in stochastic linear bandits. We show that for ensemble size , ES attains $\tilde O(d^{3…

cs.LG2026

Eluder dimension: localise it!

Alireza Bakhtiari, Alex Ayoub, Samuel Robertson +2

We establish a lower bound on the eluder dimension of generalised linear model classes, showing that standard eluder dimension-based analysis cannot lead to first-order regret boun…

cs.LG2025

High-probability zeroth-order online convex optimisation beyond Euclidean geometry

David Janz, El-Mahdi El-Mhamdi, Arya Akhavan

We study online convex optimisation with -Lipschitz losses, -regularised FTRL, and randomised two-point finite-difference gradient estimators based on cone-measure…

cs.LG2025

When and why randomised exploration works (in linear bandits)

Marc Abeille, David Janz, Ciara Pike-Burke

We provide an approach for the analysis of randomised exploration algorithms like Thompson sampling that does not rely on forced optimism or posterior inflation. With this, we demo…

cs.LG2023

Exploration via linearly perturbed loss minimisation

David Janz, Shuai Liu, Alex Ayoub +1

We introduce exploration via linear loss perturbations (EVILL), a randomised exploration method for structured stochastic bandit problems that works by solving for the minimiser of…