4 citations · 4 across the 1 of their papers we have counts for
1 paper
Anna Harutyunyan, Tim Brys, Peter Vrancx +1
Recent advances of gradient temporal-difference methods allow to learn off-policy multiple value functions in parallel with- out sacrificing convergence guarantees or computational…