2 papers
cs.CV2026
Mind the Approximation: Fisher-Weighted SVD Compression for ViTs
Moritz Thoma, Maximilian Groezinger, Maximilian Forstenhäusler +7
Model compression is key to mitigate deployment challenges of ever growing machine learning models. In this area of research, singular value decomposition (SVD)-based compression o…
cs.LG2026
HiFi-LLP: High-Fidelity, Low-Cost Latency Predictors with Confidence for Robust HW-NAS
Shambhavi Balamuthu Sampath, Behzad Shomali, Nael Fasfous +7
With deep neural networks (DNNs) increasingly deployed on edge devices, hardware (HW)-aware optimization techniques--such as HW-aware compression and HW-aware neural architecture s…