output
20022019
most citedNon-Abelian Anyons and Topological Quantum Computation

7k citations

Showing 2017Show all

60 papers · 1 filter

cs.LG201750 cited

Accelerated Gradient Descent Escapes Saddle Points Faster than Gradient Descent

Chi Jin, Praneeth Netrapalli, Michael I. Jordan

Nesterov's accelerated gradient descent (AGD), an instance of the general family of "momentum methods", provably achieves faster convergence rate than gradient descent (GD) in the…

cs.SI2017

BL-ECD: Broad Learning based Enterprise Community Detection via Hierarchical Structure Fusion

Jiawei Zhang, Limeng Cui, Philip S. Yu +1

Employees in companies can be divided into di erent communities, and those who frequently socialize with each other will be treated as close friends and are grouped in the same com…

cs.IR20172 cited

Neural Ranking Models with Multiple Document Fields

Hamed Zamani, Bhaskar Mitra, Xia Song +2

Deep neural networks have recently shown promise in the ad-hoc retrieval task. However, such models have often been based on one field of the document, for example considering docu…

cs.AI201714 cited

BBQ-Networks: Efficient Exploration in Deep Reinforcement Learning for Task-Oriented Dialogue Systems

Zachary Lipton, Xiujun Li, Jianfeng Gao +3

We present a new algorithm that significantly improves the efficiency of exploration for deep Q-learning agents in dialogue systems. Our agents explore via Thompson sampling, drawi…

stat.ML201732 cited

Z-Forcing: Training Stochastic Recurrent Networks

Anirudh Goyal, Alessandro Sordoni, Marc-Alexandre Côté +2

Many efforts have been devoted to training generative latent variable models with autoregressive decoders, such as recurrent neural networks (RNN). Stochastic recurrent models have…

cs.CL20171 cited

An Empirical Analysis of Multiple-Turn Reasoning Strategies in Reading Comprehension Tasks

Yelong Shen, Xiaodong Liu, Kevin Duh +1

Reading comprehension (RC) is a challenging task that requires synthesis of information across sentences and multiple turns of reasoning. Using a state-of-the-art RC model, we empi…