14 citations · 16 across the 2 of their papers we have counts for
3 papers
Thompson sampling for linear quadratic mean-field teams
Mukul Gagrani, Sagar Sudhakara, Aditya Mahajan +2
We consider optimal control of an unknown multi-agent linear quadratic (LQ) system where the dynamics and the cost are coupled across the agents through the mean-field (i.e., empir…
Optimal scheduling strategy for networked estimation with energy harvesting
Marcos M. Vasconcelos, Mukul Gagrani, Ashutosh Nayyar +1
Joint optimization of scheduling and estimation policies is considered for a system with two sensors and two non-collocated estimators. Each sensor produces an independent and iden…
Learning Unknown Markov Decision Processes: A Thompson Sampling Approach
Yi Ouyang, Mukul Gagrani, Ashutosh Nayyar +1
We consider the problem of learning an unknown Markov Decision Process (MDP) that is weakly communicating in the infinite horizon setting. We propose a Thompson Sampling-based rein…