3 papers
cs.CV2026
Automatic Pruning Discovery for Large Language Models
Haidong Kang, Lihong Lin, Enneng Yang +2
Large language models (LLMs) have achieved remarkable performance on a wide range of tasks, hindering real-world deployment due to their massive size. Existing pruning methods (e.g…
cs.LG2026
Unlocking the Potential of Continual Model Merging: An ODE Perspective
Lihong Lin, Haidong Kang
Continual Model Merging (CMM) enables rapid customization of foundation models by sequentially incorporating task-adapted models without repeated retraining. However, existing merg…
cs.LG2026
Revolutionizing Mixed Precision Quantization: Towards Training-free Automatic Proxy Discovery via Large Language Models
Haidong Kang, Jun Du, Lihong Lin
Mixed-Precision Quantization (MPQ) liberates Deep Neural Networks (DNNs) from the Out-Of-Memory (OOM) bottleneck and has garnered increasing research attention. However, convention…