3 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 2 cited
Actor-Critic based Improper Reinforcement Learning
Mohammadi Zaki, Avinash Mohan, Aditya Gopalan +1
We consider an improper reinforcement learning setting where a learner is given base controllers for an unknown Markov decision process, and wishes to combine them optimally to…
cs.LG2016★ 3 cited
Low-rank Bandits with Latent Mixtures
Aditya Gopalan, Odalric-Ambrym Maillard, Mohammadi Zaki
We study the task of maximizing rewards from recommending items (actions) to users sequentially interacting with a recommender system. Users are modeled as latent mixtures of C man…