4 citations · 7 across the 4 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.LG2024★ 3 cited
SGLP: A Similarity Guided Fast Layer Partition Pruning for Compressing Large Deep Models
Yuqi Li, Yao Lu, Junhao Dong +7
Layer pruning has emerged as a potent approach to remove redundant layers in the pre-trained network on the purpose of reducing network size and improve computational efficiency. H…
cs.CV2024★ 4 cited
Scaling Up Quantization-Aware Neural Architecture Search for Efficient Deep Learning on the Edge
Yao Lu, Hiram Rayo Torres Rodriguez, Sebastian Vogel +2
Neural Architecture Search (NAS) has become the de-facto approach for designing accurate and efficient networks for edge devices. Since models are typically quantized for edge depl…