PAC-Bayesian aggregation and multi-armed bandits
arXiv:1011.3396
Abstract
This habilitation thesis presents several contributions to (1) the PAC-Bayesian analysis of statistical learning, (2) the three aggregation problems: given d functions, how to predict as well as (i) the best of these d functions (model selection type aggregation), (ii) the best convex combination of these d functions, (iii) the best linear combination of these d functions, (3) the multi-armed bandit problems.
References in corpus (7)
- Consistency of the group Lasso and multiple kernel learning
- Fast learning rates for plug-in classifiers
- Empirical Bernstein Bounds and Sample Variance Penalization
- From -entropy to KL-entropy: Analysis of minimum information complexity density estimation
- Fast learning rates in statistical inference through aggregation
- PAC-Bayesian Bounds for Randomized Empirical Risk Minimizers
- High confidence estimates of the mean of heavy-tailed real random variables