From the 1 of 1.6k papers with an AI index.
24.4k citations
- University of California, BerkeleyUS97 papers
- Stanford UniversityUS79 papers
- Massachusetts Institute of TechnologyUS72 papers
- Google DeepMind (United Kingdom)GB56 papers
- Carnegie Mellon UniversityUS54 papers
- University of TorontoCA51 papers
- Princeton UniversityUS48 papers
- Cornell UniversityUS44 papers
- University of Illinois Urbana-ChampaignUS36 papers
- Columbia UniversityUS35 papers
- University of California, Santa BarbaraUS33 papers
- University of ChicagoUS30 papers
8 papers · 2 filters
BBQ-Networks: Efficient Exploration in Deep Reinforcement Learning for Task-Oriented Dialogue Systems
Zachary Lipton, Xiujun Li, Jianfeng Gao +3
We present a new algorithm that significantly improves the efficiency of exploration for deep Q-learning agents in dialogue systems. Our agents explore via Thompson sampling, drawi…
A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys +5
To achieve general intelligence, agents must learn how to interact with others in a shared environment: this is the challenge of multiagent reinforcement learning (MARL). The simpl…
Distributional Reinforcement Learning with Quantile Regression
Will Dabney, Mark Rowland, Marc G. Bellemare +1
In reinforcement learning an agent interacts with the environment by taking actions and observing the next state and reward. When sampled probabilistically, these state transitions…
Neural Program Meta-Induction
Jacob Devlin, Rudy Bunel, Rishabh Singh +2
Most recently proposed methods for Neural Program Induction work under the assumption of having a large set of input/output (I/O) examples for learning any underlying input-output…
Neural Optimizer Search with Reinforcement Learning
Irwan Bello, Barret Zoph, Vijay Vasudevan +1
We present an approach to automate the process of discovering optimization methods, with a focus on deep learning architectures. We train a Recurrent Neural Network controller to g…
Zero-Shot Task Generalization with Multi-Task Deep Reinforcement Learning
Junhyuk Oh, Satinder Singh, Honglak Lee +1
As a step towards developing zero-shot task generalization capabilities in reinforcement learning (RL), we introduce a new RL problem where the agent should learn to execute sequen…