9 citations · 11 across the 3 of their papers we have counts for
4 papers
BN-invariant sharpness regularizes the training model to better generalization
Mingyang Yi, Huishuai Zhang, Wei Chen +2
It is arguably believed that flatter minima can generalize better. However, it has been pointed out that the usual definitions of sharpness, which consider either the maxima or the…
Identifying Invariant Texture Violation for Robust Deepfake Detection
Xinwei Sun, Botong Wu, Wei Chen
Existing deepfake detection methods have reported promising in-distribution results, by accessing published large-scale dataset. However, due to the non-smooth synthesis method, th…
Dynamic of Stochastic Gradient Descent with State-Dependent Noise
Qi Meng, Shiqi Gong, Wei Chen +2
Stochastic gradient descent (SGD) and its variants are mainstream methods to train deep neural networks. Since neural networks are non-convex, more and more works study the dynamic…
Interpreting Basis Path Set in Neural Networks
Juanping Zhu, Qi Meng, Wei Chen +1
Based on basis path set, G-SGD algorithm significantly outperforms conventional SGD algorithm in optimizing neural networks. However, how the inner mechanism of basis paths work re…