3 papers
cs.LG2020
Policy Gradient using Weak Derivatives for Reinforcement Learning
Sujay Bhatt, Alec Koppel, Vikram Krishnamurthy
This paper considers policy search in continuous state-action reinforcement learning problems. Typically, one computes search directions using a classic expression for the policy g…
cs.SI2018
Adaptive Polling in Hierarchical Social Networks using Blackwell Dominance
Sujay Bhatt, Vikram Krishnamurthy
Consider a population of individuals that observe an underlying state of nature that evolves over time. The population is classified into different levels depending on the hierarch…
math.OC2017
Controlled Information Fusion with Risk-Averse CVaR Social Sensors
Sujay Bhatt, Vikram Krishnamurthy
Consider a multi-agent network comprised of risk averse social sensors and a controller that jointly seek to estimate an unknown state of nature, given noisy measurements. The netw…