1.1k citations · 2.9k across the 7 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2016★ 1.1k cited
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt +2
Despite their massive size, successful deep artificial neural networks can exhibit a remarkably small difference between training and test performance. Conventional wisdom attribut…
cs.LG2016★ 16 cited
Can Active Memory Replace Attention?
Łukasz Kaiser, Samy Bengio
Several mechanisms to focus attention of a neural network on selected parts of its input or memory have been used successfully in deep learning models in recent years. Attention ha…
cs.LG2016★ 88 cited
Reward Augmented Maximum Likelihood for Neural Structured Prediction
Mohammad Norouzi, Samy Bengio, Zhifeng Chen +4
A key problem in structured output prediction is direct optimization of the task reward function that matters for test evaluation. This paper presents a simple and computationally…