3 papers
cs.LG2026
Contextual Slate GLM Bandits with Limited Adaptivity
Tanmay Goyal, Sukruta Prakash Midigeshi, Gaurav Sinha
We investigate the contextual slate bandit problem with generalized linear rewards under limited adaptivity. At each round, the learner is presented with sets of items, where e…
cs.LG2026
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
Tanmay Goyal, Gaurav Sinha
We study the Logistic Contextual Slate Bandit problem, where, at each round, an agent selects a slate of items from an exponentially large set (of size ) of candidat…
cs.LG2025
Achieving Limited Adaptivity for Multinomial Logistic Bandits
Sukruta Prakash Midigeshi, Tanmay Goyal, Gaurav Sinha
Multinomial Logistic Bandits have recently attracted much attention due to their ability to model problems with multiple outcomes. In this setting, each decision is associated with…