1 citations · 2 across the 3 of their papers we have counts for
4 papers
Model-based micro-data reinforcement learning: what are the crucial model properties and which model to choose?
Balázs Kégl, Gabriel Hurtado, Albert Thomas
We contribute to micro-data model-based reinforcement learning (MBRL) by rigorously comparing popular generative models using a fixed (random shooting) control agent. We find that…
Refined bounds for randomized experimental design
Geovani Rizk, Igor Colin, Albert Thomas +1
Experimental design is an approach for selecting samples among a given set so as to obtain the best estimator for a given criterion. In the context of linear regression, several op…
Best Arm Identification in Graphical Bilinear Bandits
Geovani Rizk, Albert Thomas, Igor Colin +2
We introduce a new graphical bilinear bandit problem where a learner (or a \emph{central entity}) allocates arms to the nodes of a graph and observes for each edge a noisy bilinear…
Parallel Contextual Bandits in Wireless Handover Optimization
Igor Colin, Albert Thomas, Moez Draief
As cellular networks become denser, a scalable and dynamic tuning of wireless base station parameters can only be achieved through automated optimization. Although the contextual b…