7k citations
- University of California, Santa BarbaraUS105 papers
- University of California, BerkeleyUS38 papers
- ETH ZurichCH37 papers
- University of Maryland, College ParkUS32 papers
- Stanford UniversityUS31 papers
- California Institute of TechnologyUS25 papers
- Microsoft Research (United Kingdom)GB25 papers
- Princeton UniversityUS21 papers
- Cornell UniversityUS20 papers
- Carnegie Mellon UniversityUS19 papers
- University of California, Los AngelesUS18 papers
- Microsoft Research New York City (United States)17 papers
60 papers · 1 filter
Accelerated Gradient Descent Escapes Saddle Points Faster than Gradient Descent
Chi Jin, Praneeth Netrapalli, Michael I. Jordan
Nesterov's accelerated gradient descent (AGD), an instance of the general family of "momentum methods", provably achieves faster convergence rate than gradient descent (GD) in the…
BL-ECD: Broad Learning based Enterprise Community Detection via Hierarchical Structure Fusion
Jiawei Zhang, Limeng Cui, Philip S. Yu +1
Employees in companies can be divided into di erent communities, and those who frequently socialize with each other will be treated as close friends and are grouped in the same com…
Neural Ranking Models with Multiple Document Fields
Hamed Zamani, Bhaskar Mitra, Xia Song +2
Deep neural networks have recently shown promise in the ad-hoc retrieval task. However, such models have often been based on one field of the document, for example considering docu…
BBQ-Networks: Efficient Exploration in Deep Reinforcement Learning for Task-Oriented Dialogue Systems
Zachary Lipton, Xiujun Li, Jianfeng Gao +3
We present a new algorithm that significantly improves the efficiency of exploration for deep Q-learning agents in dialogue systems. Our agents explore via Thompson sampling, drawi…
Z-Forcing: Training Stochastic Recurrent Networks
Anirudh Goyal, Alessandro Sordoni, Marc-Alexandre Côté +2
Many efforts have been devoted to training generative latent variable models with autoregressive decoders, such as recurrent neural networks (RNN). Stochastic recurrent models have…
An Empirical Analysis of Multiple-Turn Reasoning Strategies in Reading Comprehension Tasks
Yelong Shen, Xiaodong Liu, Kevin Duh +1
Reading comprehension (RC) is a challenging task that requires synthesis of information across sentences and multiple turns of reasoning. Using a state-of-the-art RC model, we empi…