64 citations · 168 across the 23 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2018
Addressing Function Approximation Error in Actor-Critic Methods
Scott Fujimoto, Herke van Hoof, David Meger
In value-based reinforcement learning methods such as deep Q-learning, function approximation errors are known to lead to overestimated value estimates and suboptimal policies. We…
cs.AI2017★ 17 cited
Benchmark Environments for Multitask Learning in Continuous Domains
Peter Henderson, Wei-Di Chang, Florian Shkurti +3
As demand drives systems to generalize to various domains and problems, the study of multitask, transfer and lifelong learning has become an increasingly important pursuit. In disc…