2 papers
cs.CV2019
DeepSquare: Boosting the Learning Power of Deep Convolutional Neural Networks with Elementwise Square Operators
Sheng Chen, Xu Wang, Chao Chen +3
Modern neural network modules which can significantly enhance the learning power usually add too much computational complexity to the original neural networks. In this paper, we pu…
cs.LG2018
EA-CG: An Approximate Second-Order Method for Training Fully-Connected Neural Networks
Sheng-Wei Chen, Chun-Nan Chou, Edward Y. Chang
For training fully-connected neural networks (FCNNs), we propose a practical approximate second-order method including: 1) an approximation of the Hessian matrix and 2) a conjugate…