14 citations · 23 across the 3 of their papers we have counts for
1 paper · 1 filter
Qi Meng, Shuxin Zheng, Huishuai Zhang +3
It is well known that neural networks with rectified linear units (ReLU) activation functions are positively scale-invariant. Conventional algorithms like stochastic gradient desce…