6.7k citations · 6.7k across the 1 of their papers we have counts for
1 paper
Geoffrey E. Hinton, Nitish Srivastava, Alex Krizhevsky +2
When a large feedforward neural network is trained on a small training set, it typically performs poorly on held-out test data. This "overfitting" is greatly reduced by randomly om…