2 citations · 3 across the 3 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024★ 1 cited
ColA: Collaborative Adaptation with Gradient Learning
Enmao Diao, Qi Le, Suya Wu +4
A primary function of back-propagation is to compute both the gradient of hidden representations and parameters for optimization with gradient descent. Training large models requir…
cs.LG2022
On The Energy Statistics of Feature Maps in Pruning of Neural Networks with Skip-Connections
Mohammadreza Soltani, Suya Wu, Yuerong Li +2
We propose a new structured pruning framework for compressing Deep Neural Networks (DNNs) with skip connections, based on measuring the statistical dependency of hidden layers and…