13 citations · 22 across the 3 of their papers we have counts for
3 papers
cs.CL2022★ 2 cited
One-Shot Learning from a Demonstration with Hierarchical Latent Language
Nathaniel Weir, Xingdi Yuan, Marc-Alexandre Côté +5
Humans have the capability, aided by the expressive compositionality of their language, to learn quickly by demonstration. They are able to describe unseen task-performing procedur…
cs.LG2016★ 7 cited
Separation of Concerns in Reinforcement Learning
Harm van Seijen, Mehdi Fatemi, Joshua Romoff +1
In this paper, we propose a framework for solving a single-agent task by using multiple agents, each focusing on different aspects of the task. This approach has two main advantage…
cs.AI2016★ 13 cited
Effective Multi-step Temporal-Difference Learning for Non-Linear Function Approximation
Harm van Seijen
Multi-step temporal-difference (TD) learning, where the update targets contain information from multiple time steps ahead, is one of the most popular forms of TD learning for linea…