2 papers
cs.LG2026
Learning What to Recommend: Minimax Optimal Simple Regret in Logistic Bandits
Shuai Liu, Alireza Bakhtiari, Alex Ayoub +2
We study stochastic logistic bandits with -dimensional action features under the simple-regret objective, where a learner uses rounds of exploration to output a single final…
cs.LG2026
Eluder dimension: localise it!
Alireza Bakhtiari, Alex Ayoub, Samuel Robertson +2
We establish a lower bound on the eluder dimension of generalised linear model classes, showing that standard eluder dimension-based analysis cannot lead to first-order regret boun…