2 papers
cs.LG2026
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
Ashutosh Soni, Peizhong Ju, Atilla Eryilmaz +1
We study the stochastic multi-armed bandit (MAB) problem where an underlying network structure enables side-observations across related actions. We use a bipartite graph to link ac…
cs.LG2025
BeST -- A Novel Source Selection Metric for Transfer Learning
Ashutosh Soni, Peizhong Ju, Atilla Eryilmaz +1
One of the most fundamental, and yet relatively less explored, goals in transfer learning is the efficient means of selecting top candidates from a large number of previously train…