1 citations · 1 across the 4 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Neural Precision Polarization: Simplifying Neural Network Inference with Dual-Level Precision
Dinithi Jayasuriya, Nastaran Darabi, Maeesha Binte Hashem +1
We introduce a precision polarization scheme for DNN inference that utilizes only very low and very high precision levels, assigning low precision to the majority of network weight…
cs.LG2023★ 1 cited
Towards Model-Size Agnostic, Compute-Free, Memorization-based Inference of Deep Learning
Davide Giacomini, Maeesha Binte Hashem, Jeremiah Suarez +2
The rapid advancement of deep neural networks has significantly improved various tasks, such as image and speech recognition. However, as the complexity of these models increases,…