16 citations · 17 across the 5 of their papers we have counts for
Showing 2019 · cs.LGShow all
2 papers · 2 filters
cs.LG2019
Reinforcement Learning for Joint Optimization of Multiple Rewards
Mridul Agarwal, Vaneet Aggarwal
Finding optimal policies which maximize long term rewards of Markov Decision Processes requires the use of dynamic programming and backward induction to solve the Bellman optimalit…
cs.LG2019
Reinforcement Learning for Mean Field Game
Mridul Agarwal, Vaneet Aggarwal, Arnob Ghosh +1
Stochastic games provide a framework for interactions among multiple agents and enable a myriad of applications. In these games, agents decide on actions simultaneously, the state…