209 citations · 259 across the 2 of their papers we have counts for
2 papers
cs.LG2013★ 50 cited
Big Neural Networks Waste Capacity
Yann N. Dauphin, Yoshua Bengio
This article exposes the failure of some big neural networks to leverage added capacity to reduce underfitting. Past research suggest diminishing returns when increasing the size o…
cs.LG2012★ 209 cited
Better Mixing via Deep Representations
Yoshua Bengio, Grégoire Mesnil, Yann Dauphin +1
It has previously been hypothesized, and supported with some experimental evidence, that deeper representations, when well trained, tend to do a better job at disentangling the und…