Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
The Structural Scalpel: Automated Contiguous Layer Pruning for Large Language Models
Yao Lu, Yuqi Li, Wenbin Xie +4
Although large language models (LLMs) have achieved revolutionary breakthroughs in many fields, their large model size and high computational cost pose significant challenges for p…
cs.LG2025
LoRALib: A Standardized Benchmark for Evaluating LoRA-MoE Methods
Shaoheng Wang, Yao Lu, Yuqi Li +5
As a parameter efficient fine-tuning (PEFT) method, low-rank adaptation (LoRA) can save significant costs in storage and computing, but its strong adaptability to a single task is…