2 papers
cs.LG2025
Exploiting Adjacent Similarity in Multi-Armed Bandit Tasks via Transfer of Reward Samples
NR Rahul, Vaibhav Katewa
We consider a sequential multi-task problem, where each task is modeled as the stochastic multi-armed bandit with K arms. We assume the bandit tasks are adjacently similar in the s…
cs.LG2024
Transfer in Sequential Multi-armed Bandits via Reward Samples
Rahul N R, Vaibhav Katewa
We consider a sequential stochastic multi-armed bandit problem where the agent interacts with bandit over multiple episodes. The reward distribution of the arms remain constant thr…