37 citations · 62 across the 15 of their papers we have counts for
3 papers · 1 filter
Wasserstein-Splitting Gaussian Process Regression for Heterogeneous Online Bayesian Inference
Michael E. Kepler, Alec Koppel, Amrit Singh Bedi +1
Gaussian processes (GPs) are a well-known nonparametric Bayesian inference technique, but they suffer from scalability problems for large sample sizes, and their performance can de…
MARL with General Utilities via Decentralized Shadow Reward Actor-Critic
Junyu Zhang, Amrit Singh Bedi, Mengdi Wang +1
We posit a new mechanism for cooperation in multi-agent reinforcement learning (MARL) based upon any nonlinear function of the team's long-term state-action occupancy measure, i.e.…
Cautious Reinforcement Learning via Distributional Risk in the Dual Domain
Junyu Zhang, Amrit Singh Bedi, Mengdi Wang +1
We study the estimation of risk-sensitive policies in reinforcement learning problems defined by a Markov Decision Process (MDPs) whose state and action spaces are countably finite…