Showing cs.LGShow all
2 papers · 1 filter
cs.LG2019
A Convergence Result for Regularized Actor-Critic Methods
Wesley Suttle, Zhuoran Yang, Kaiqing Zhang +1
In this paper, we present a probability one convergence proof, under suitable conditions, of a certain class of actor-critic algorithms for finding approximate solutions to entropy…
cs.LG2019
A Multi-Agent Off-Policy Actor-Critic Algorithm for Distributed Reinforcement Learning
Wesley Suttle, Zhuoran Yang, Kaiqing Zhang +3
This paper extends off-policy reinforcement learning to the multi-agent case in which a set of networked agents communicating with their neighbors according to a time-varying graph…