2 citations · 6 across the 14 of their papers we have counts for
1 paper · 1 filter
Donghwan Lee, Jianghai Hu
The goal of this paper is to study a distributed version of the gradient temporal-difference (GTD) learning algorithm for a class of multi-agent Markov decision processes (MDPs). T…