2 papers
cs.LG2026
Bayesian Online Model Selection
Aida Afshar, Yuke Zhang, Aldo Pacchiano
Online model selection in Bayesian bandits raises a fundamental exploration challenge: When an environment instance is sampled from a prior distribution, how can we design an adapt…
cs.LG2025
Improved Training Mechanism for Reinforcement Learning via Online Model Selection
Aida Afshar, Aldo Pacchiano
We study the problem of online model selection in reinforcement learning, where the selector has access to a class of reinforcement learning agents and learns to adaptively select…