4 citations · 10 across the 6 of their papers we have counts for
3 papers · 1 filter
BabyAI 1.1
David Yu-Tung Hui, Maxime Chevalier-Boisvert, Dzmitry Bahdanau +1
The BabyAI platform is designed to measure the sample efficiency of training an agent to follow grounded-language instructions. BabyAI 1.0 presents baseline results of an agent tra…
Option-Critic in Cooperative Multi-agent Systems
Jhelum Chakravorty, Nadeem Ward, Julien Roy +4
In this paper, we investigate learning temporal abstractions in cooperative multi-agent systems, using the options framework (Sutton et al, 1999). First, we address the planning pr…
BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning
Maxime Chevalier-Boisvert, Dzmitry Bahdanau, Salem Lahlou +4
Allowing humans to interactively train artificial agents to understand language instructions is desirable for both practical and scientific reasons, but given the poor data efficie…