Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
On the necessity of adaptive regularisation:Optimal anytime online learning on -balls
Emmeran Johnson, David Martínez-Rubio, Ciara Pike-Burke +1
We study online convex optimization on -balls in for . While always sub-linear, the optimal regret exhibits a shift between the high-dimensional setti…
cs.LG2018
Decentralized Cooperative Stochastic Bandits
David Martínez-Rubio, Varun Kanade, Patrick Rebeschini
We study a decentralized cooperative stochastic multi-armed bandit problem with arms on a network of agents. In our model, the reward distribution of each arm is the same f…