output
20052015
most citedBatch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift

24.4k citations

Showing 2014 · cs.LGShow all

5 papers · 2 filters

cs.LG20143 cited

ACCAMS: Additive Co-Clustering to Approximate Matrices Succinctly

Alex Beutel, Amr Ahmed, Alexander J. Smola

Matrix completion and approximation are popular tools to capture a user's preferences for recommendation and to approximate missing data. Instead of using low-rank factorization we…

cs.LG2014699 cited

Multiple Object Recognition with Visual Attention

Jimmy Ba, Volodymyr Mnih, Koray Kavukcuoglu

We present an attention-based model for recognizing multiple objects in images. The proposed model is a deep recurrent neural network trained with reinforcement learning to attend…

cs.LG201493 cited

Move Evaluation in Go Using Deep Convolutional Neural Networks

Chris J. Maddison, Aja Huang, Ilya Sutskever +1

The game of Go is more challenging than other board games, due to the difficulty of constructing a position or move evaluation function. In this paper we investigate whether deep c…

cs.LG201411 cited

Skip-gram Language Modeling Using Sparse Non-negative Matrix Probability Estimation

Noam Shazeer, Joris Pelemans, Ciprian Chelba

We present a novel family of language model (LM) estimation techniques named Sparse Non-negative Matrix (SNM) estimation. A first set of experiments empirically evaluating it on th…

cs.LG201418 cited

Differentially- and non-differentially-private random decision trees

Mariusz Bojarski, Anna Choromanska, Krzysztof Choromanski +1

We consider supervised learning with random decision trees, where the tree construction is completely random. The method is popularly used and works well in practice despite the si…