Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
DLR: Zero-Inference-Cost Latent Residuals for Low-Rank Pre-Training
Dong Wang, Wenwu Tang, Yun Cheng +1
Large language models have driven recent progress in language and multimodal AI, yet pre-training them at scale is prohibitively expensive. Low-rank pre-training, which factorizes…
cs.LG2026
Cut Less, Fold More: Model Compression through the Lens of Projection Geometry
Olga Saukh, Dong Wang, Haris Å ikiÄ +2
Compressing neural networks without retraining is vital for deployment at scale. We study calibration-free compression through the lens of projection geometry: structured pruning i…