2 papers
cs.LG2013
Reinforcement Learning for Matrix Computations: PageRank as an Example
Vivek S. Borkar, Adwaitvedant S. Mathkar
Reinforcement learning has gained wide popularity as a technique for simulation-driven approximate dynamic programming. A less known aspect is that the very reasons that make it ef…
cs.DC2013
Distributed Reinforcement Learning via Gossip
Adwaitvedant S. Mathkar, Vivek S. Borkar
We consider the classical TD(0) algorithm implemented on a network of agents wherein the agents also incorporate the updates received from neighboring agents using a gossip-like me…