2 papers
cs.LG2024
Meta Learning in Bandits within Shared Affine Subspaces
Steven Bilaj, Sofien Dhouib, Setareh Maghsudi
We study the problem of meta-learning several contextual stochastic bandits tasks by leveraging their concentration around a low-dimensional affine subspace, which we learn via onl…
cs.LG2023
Piecewise-Stationary Combinatorial Semi-Bandit with Causally Related Rewards
Behzad Nourani-Koliji, Steven Bilaj, Amir Rezaei Balef +1
We study the piecewise stationary combinatorial semi-bandit problem with causally related rewards. In our nonstationary environment, variations in the base arms' distributions, cau…