2 papers
cs.LG2019
Forget the Learning Rate, Decay Loss
Jiakai Wei
In the usual deep neural network optimization process, the learning rate is the most important hyper parameter, which greatly affects the final convergence effect. The purpose of l…
cs.LG2018
Fast, Better Training Trick -- Random Gradient
Jiakai Wei
In this paper, we will show an unprecedented method to accelerate training and improve performance, which called random gradient (RG). This method can be easier to the training of…