37 citations · 87 across the 10 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2021★ 5 cited
MARL with General Utilities via Decentralized Shadow Reward Actor-Critic
Junyu Zhang, Amrit Singh Bedi, Mengdi Wang +1
We posit a new mechanism for cooperation in multi-agent reinforcement learning (MARL) based upon any nonlinear function of the team's long-term state-action occupancy measure, i.e.…
stat.ML2020★ 10 cited
Cautious Reinforcement Learning via Distributional Risk in the Dual Domain
Junyu Zhang, Amrit Singh Bedi, Mengdi Wang +1
We study the estimation of risk-sensitive policies in reinforcement learning problems defined by a Markov Decision Process (MDPs) whose state and action spaces are countably finite…