1k citations · 1.5k across the 7 of their papers we have counts for
Showing 2020Show all
2 papers · 1 filter
cs.AI2020★ 6 cited
Relative Variational Intrinsic Control
Kate Baumli, David Warde-Farley, Steven Hansen +1
In the absence of external rewards, agents can still learn useful behaviors by identifying and mastering a set of diverse skills within their environment. Existing skill learning m…
cs.LG2020★ 30 cited
Q-Learning in enormous action spaces via amortized approximate maximization
Tom Van de Wiele, David Warde-Farley, Andriy Mnih +1
Applying Q-learning to high-dimensional or continuous action spaces can be difficult due to the required maximization over the set of possible actions. Motivated by techniques from…